RISE: Recursive Improvement via Self-Extrapolating Policy Distillation Paper • 2609.05295 • Published 3 days ago • 2
Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference Paper • 2609.05275 • Published 3 days ago • 3
WorldSculpt: Generating Compositional Worlds from Grounded Videos Paper • 2609.05416 • Published 3 days ago • 2
Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments Paper • 2609.04148 • Published 4 days ago • 272
Post-Training Language Models for Gold-Medal Performance in Coding Competitions Paper • 2609.02849 • Published 5 days ago • 9
Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement Paper • 2609.01481 • Published 6 days ago • 17
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper • 2609.00111 • Published 7 days ago • 374
Matrix-Game 3.5: Enhancing Real-Time Streaming Interactive World Models with Patch Memory Paper • 2608.29910 • Published 8 days ago • 18
PaperBanana-Interact: Scientific Diagram Refinement with Multi-Turn Human Feedback Paper • 2608.30241 • Published 7 days ago • 12
On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Paper • 2608.30320 • Published 7 days ago • 53
Scaffolding Foundation Models into Physical-World Agents Pushes the Frontier of Long-Horizon Navigation Paper • 2608.30396 • Published 7 days ago • 10
BLARM: Animating 3D Objects from Video via Blending Latent Rigid Motion Primitives Paper • 2608.31113 • Published 7 days ago • 4
Rubric-to-Code Credit Assignment for Reinforcement Learning Paper • 2608.27906 • Published 10 days ago • 7