DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence Paper • 2606.19348 • Published Apr 26 • 19
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization Paper • 2607.25659 • Published 4 days ago • 80
K-EXAONE-2.0 Collection Journey to global frontier-scale foundation models • 4 items • Updated 1 day ago • 54
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 3 days ago • 123
CodeNib: A Multi-View Data System for Serving Repository Context to Coding Agents Paper • 2607.25431 • Published 4 days ago • 98
HiPO Collection Adaptive reasoning LLMs based on the HiPO framework, featuring dynamic “think-on / think-off” control for efficient reasoning. • 2 items • Updated Nov 3, 2025 • 5
Qwen3.5-Claude-Fable-5 Collection Our series of Qwen3.5 finetunes on Claude-Fable-5 outputs • 2 items • Updated Jun 19 • 9
Qwythos v1 Collection A collection of our uncensored Qwen Claude Mythos fine tunes • 2 items • Updated 21 days ago • 36
Qwythos v2 Collection A collection of our uncensored Qwen Claude Mythos fine tunes • 2 items • Updated 21 days ago • 11
Very Large GGUFs Collection GGUF quantized versions of very large models - over 100B parameters • 87 items • Updated 4 days ago • 10
view article Article The Open Source Community is backing OpenEnv for Agentic RL +18 burtenshaw, spisakjo, lysandre, darktex, willcb, qjoy, pawalt, cwing-nv, danielhanchen, andrewzhou, thegovind, shimmyshimmer, Hamid-Nazeri, Sanyam, zkwentz, emre0, lewtun, sergiopaniego, banghua, unseenmars • Jun 8 • 109
RoboBrain-Dex Collection Dexterous VLA utilizing human ego data training • 2 items • Updated Mar 13 • 5