astune/text_info_trending_youtube_videos_2019-04-15_to_2020-04-15 Viewer • Updated May 3 • 10.6k • 97 • 6
The Past Frames the Future: Memory for Autoregressive Video Generation Paper • 2609.28466 • Published 7 days ago • 64
InternW0: A Foundational Physical World Model for Efficient Real-World Interactions Paper • 2609.27656 • Published 7 days ago • 14
Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 7 days ago • 51
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 9 days ago • 156
Why Do Video Diffusion Models Violate Physics? Unveiling the Flaws in Attention Mechanisms Paper • 2609.23658 • Published 10 days ago • 30
All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts Paper • 2609.24058 • Published 9 days ago • 55
Flash-dLLM: IO-Aware KV Caching and Parallel Decoding for Fast, Memory-Efficient Diffusion LLMs Paper • 2609.26796 • Published 8 days ago • 36
Think Like a World Model, Act Like a VLA: Distilling World-Model Representations into Compact Robot Policies Paper • 2609.24682 • Published 9 days ago • 12
EDGEGEN: Improving Tool-Calling Agents Beyond Happy Paths with Synthetic Edge Case Generation Paper • 2609.24115 • Published 9 days ago • 5
FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations Paper • 2609.20817 • Published 13 days ago • 40
VākQA: A Benchmark and Evaluation Study for Telugu Spoken Factoid Question Answering Paper • 2609.19879 • Published 13 days ago • 33