Keep It InMind: Benchmarking the Implicit-Association Blind Spot in Agent Memory Paper • 2607.24368 • Published 7 days ago • 31
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Paper • 2607.24904 • Published 7 days ago • 32
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 4 days ago • 48
HumanCLAW: Can Vision-Language Models Act Through a Body? Paper • 2607.27180 • Published 5 days ago • 71
Running on Zero Agents Featured 88 Mage-VL 🎞 88 Codec-native video & image understanding with Mage-VL 4B
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Paper • 2607.17423 • Published 15 days ago • 166
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published 8 days ago • 123
GNM Head: A Generative aNthropometric Model of the human head Paper • 2607.23687 • Published 8 days ago • 5