StarHarness: Evolving Harnesses with Stratified Search for Enterprise Environments Paper • 2608.24804 • Published 13 days ago • 40
A Modular Agent for Reliable and Auditable Spatial Relation Verification in CT Scans Paper • 2608.21140 • Published 17 days ago • 4
ZimaBlue: Evolving Generalizable World Action Models through Scalable Video Pre-training Paper • 2609.00188 • Published 7 days ago • 51
ajtamayoh/NLP-CIC-WFU_Clinical_Cases_NER_Paragraph_Tokenized_mBERT_cased_fine_tuned Token Classification • Updated Jun 9, 2022 • 21 • 1
StepGuard: Learning Step-Level Guardrails with Scalable Supervision and Safety-Utility Balancing Paper • 2608.24777 • Published 13 days ago • 16
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 12 days ago • 196
UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City Paper • 2608.27456 • Published 11 days ago • 112
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 18 days ago • 111
AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces Paper • 2608.23041 • Published 14 days ago • 64
CAFE: Self-Improving Search Agents Need Co-Evolving Feedback Paper • 2608.24794 • Published 13 days ago • 6