ZimaBlue: Evolving Generalizable World Action Models through Scalable Video Pre-training Paper • 2609.00188 • Published 5 days ago • 46
StepGuard: Learning Step-Level Guardrails with Scalable Supervision and Safety-Utility Balancing Paper • 2608.24777 • Published 11 days ago • 16
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 10 days ago • 196
UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City Paper • 2608.27456 • Published 9 days ago • 111
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs Paper • 2608.20492 • Published 16 days ago • 111
AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces Paper • 2608.23041 • Published 12 days ago • 64
CAFE: Self-Improving Search Agents Need Co-Evolving Feedback Paper • 2608.24794 • Published 11 days ago • 6