SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 10 days ago • 268
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 6 days ago • 182
Drift-Constrained Optimization: Only Direction Matters in Fine-Tuning Instruct Models Paper • 2609.13680 • Published 8 days ago • 9
ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals Paper • 2609.16816 • Published 5 days ago • 7
ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents Paper • 2609.17523 • Published 5 days ago • 26
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 13 days ago • 341
MInTRL: Off-policy Intervention can boost On-policy RL Paper • 2609.12419 • Published 9 days ago • 11
HazardAuditor: From Executable Threats to Safer Computer-Use Agents Paper • 2609.15134 • Published 6 days ago • 16
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 6 days ago • 231
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 9 days ago • 252