Running 602 Scaling test-time compute π 602 Boost LLM answers with flexible testβtime search strategies
Running 4.01k The Ultra-Scale Playbook π 4.01k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.29k The Smol Training Playbook π 3.29k The secrets to building world-class LLMs
deepseek-ai/DeepSeek-V3.2-Speciale Text Generation β’ 685B β’ Updated Dec 1, 2025 β’ 5.73k β’ 724
Running on Zero Agents Featured 453 DeepSeek OCR Demo π 453 An interactive demo for the DeepSeek-OCR model.