Running 3.98k The Ultra-Scale Playbook 🌌 3.98k The ultimate guide to training LLM on large GPU Clusters
Running Featured 1.41k FineWeb: decanting the web for the finest text data at scale 🍷 1.41k Explore and download the FineWeb web‑scale text dataset
nvidia/Llama-3_3-Nemotron-Super-49B-v1_5-NVFP4 Text Generation • 26B • Updated Nov 27, 2025 • 26.2k • 21
Running on CPU Upgrade Featured 3.27k The Smol Training Playbook 📚 3.27k The secrets to building world-class LLMs
Vikhrmodels/Vikhr-Qwen-2.5-0.5B-instruct-GGUF Text Generation • 0.5B • Updated Oct 6, 2024 • 504 • 10