Note: The solution may not be in `solution` or `answer` columns, but inside /boxed/{ANSWER}
🔄 In a Training Loop
Gurvaah Singh
ReallyFloppyPenguin
AI & ML interests
AI, GGUFing AI, AI, Running AI, Thinking about AI, and so on
Recent Activity
liked a model 2 days ago
nineninesix/diamond-1.0 liked a Space 3 days ago
hugging-apps/rynnbrain1-1-2b-demo liked a model 3 days ago
moonshotai/Kimi-K2.7-CodeOrganizations
Datasets That Kill
Sikh Models
-
HuggingFaceTB/SmolLM3-3B
Text Generation • 3B • Updated • 782k • 985 -
Qwen/Qwen3-4B
Text Generation • 4B • Updated • 4.8M • • 662 -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 8.25M • • 6.39k -
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 5.69M • 2.73k
GGUFs
Interesting Papers
-
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 108 -
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
System Prompt Optimization with Meta-Learning
Paper • 2505.09666 • Published • 72 -
Visual Planning: Let's Think Only with Images
Paper • 2505.11409 • Published • 57
MathRL
Note: The solution may not be in `solution` or `answer` columns, but inside /boxed/{ANSWER}
Datasets That Kill
Free AI!!!
Sikh Models
-
HuggingFaceTB/SmolLM3-3B
Text Generation • 3B • Updated • 782k • 985 -
Qwen/Qwen3-4B
Text Generation • 4B • Updated • 4.8M • • 662 -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 8.25M • • 6.39k -
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 5.69M • 2.73k
Revolutionary Models
GGUFs
Ultra Cool Models
Interesting Papers
-
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 108 -
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
System Prompt Optimization with Meta-Learning
Paper • 2505.09666 • Published • 72 -
Visual Planning: Let's Think Only with Images
Paper • 2505.11409 • Published • 57