SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 10 days ago • 268
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 6 days ago • 191
Drift-Constrained Optimization: Only Direction Matters in Fine-Tuning Instruct Models Paper • 2609.13680 • Published 8 days ago • 10
ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals Paper • 2609.16816 • Published 5 days ago • 8
ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents Paper • 2609.17523 • Published 5 days ago • 27
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 13 days ago • 347
MInTRL: Off-policy Intervention can boost On-policy RL Paper • 2609.12419 • Published 9 days ago • 12
HazardAuditor: From Executable Threats to Safer Computer-Use Agents Paper • 2609.15134 • Published 6 days ago • 18
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 6 days ago • 236
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 9 days ago • 254