Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training Paper • 2609.07108 • Published 3 days ago • 23
BeaconKV: Key-Value Cache Compression Guided by Beacon Queries for Efficient Large Reasoning Model Inference Paper • 2609.04971 • Published 6 days ago • 29
VeriPhy: Agentic Physical Reasoning for World Model Evaluation and Refinement Paper • 2609.03153 • Published 8 days ago • 13