TeichAI/GLM-4.7-Flash-Claude-Opus-4.5-High-Reasoning-Distill-GGUF 30B • Updated Feb 22 • 3.94k • 512
view article Article Continuous batching from first principles +1 ror, ArthurZ, mcpotato • Nov 25, 2025 • 443
Running 4.03k The Ultra-Scale Playbook 🌌 4.03k The ultimate guide to training LLM on large GPU Clusters