Note: The solution may not be in `solution` or `answer` columns, but inside /boxed/{ANSWER}
🔄 In a Training Loop
Gurvaah Singh
ReallyFloppyPenguin
AI & ML interests
AI, GGUFing AI, AI, Running AI, Thinking about AI, and so on
Recent Activity
liked a model 3 days ago
moonshotai/Kimi-K3 liked a model 8 days ago
nineninesix/diamond-1.0 liked a Space 9 days ago
hugging-apps/rynnbrain1-1-2b-demoOrganizations
Datasets That Kill
Sikh Models
-
HuggingFaceTB/SmolLM3-3B
Text Generation • 3B • Updated • 865k • 987 -
Qwen/Qwen3-4B
Text Generation • 4B • Updated • 4.49M • • 668 -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 7.95M • • 6.44k -
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 5.18M • 2.74k
GGUFs
Interesting Papers
-
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 108 -
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
System Prompt Optimization with Meta-Learning
Paper • 2505.09666 • Published • 72 -
Visual Planning: Let's Think Only with Images
Paper • 2505.11409 • Published • 57
MathRL
Note: The solution may not be in `solution` or `answer` columns, but inside /boxed/{ANSWER}
Datasets That Kill
Free AI!!!
Sikh Models
-
HuggingFaceTB/SmolLM3-3B
Text Generation • 3B • Updated • 865k • 987 -
Qwen/Qwen3-4B
Text Generation • 4B • Updated • 4.49M • • 668 -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 7.95M • • 6.44k -
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 5.18M • 2.74k
Revolutionary Models
GGUFs
Ultra Cool Models
Interesting Papers
-
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 108 -
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
System Prompt Optimization with Meta-Learning
Paper • 2505.09666 • Published • 72 -
Visual Planning: Let's Think Only with Images
Paper • 2505.11409 • Published • 57