Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems Paper • 2412.09413 • Published Dec 12, 2024 • 1 • 1
Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning Paper • 2412.09078 • Published Dec 12, 2024 • 1
OVEarth-Bench: Evaluating Category Breadth and Query Diversity for Open-Vocabulary Earth Observation Paper • 2607.27278 • Published 6 days ago • 13 • 2
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine Paper • 2607.28625 • Published 5 days ago • 36 • 1
$β$-OPSD: Deriving with Policy Optimization, Training with Self-Distillation Paper • 2607.28582 • Published 5 days ago • 21 • 2
$Σ$-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems Paper • 2607.27958 • Published 5 days ago • 13 • 2
Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes Paper • 2607.26627 • Published 6 days ago • 4 • 2
See2Think: Do Multimodal Models Really Use Intermediate Visual States? Paper • 2607.26769 • Published 6 days ago • 23 • 2
SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them Paper • 2607.27703 • Published 5 days ago • 21 • 2
RefCaptioner: Multi-Reference Image-Grounded Video Captioning Paper • 2607.28509 • Published 5 days ago • 25 • 2
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 5 days ago • 48 • 2
Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale Paper • 2607.28074 • Published 5 days ago • 10 • 2
Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation Paper • 2607.27816 • Published 5 days ago • 32 • 2
AMRD: Adaptive Multi-Teacher Relational Distillation for Lightweight Speech Emotion Recognition Paper • 2607.25289 • Published 7 days ago • 2