Running 602 Scaling test-time compute π 602 Boost LLM answers with flexible testβtime search strategies
Running 4.05k The Ultra-Scale Playbook π 4.05k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook π 3.31k The secrets to building world-class LLMs
deepseek-ai/DeepSeek-V3.2-Speciale Text Generation β’ 685B β’ Updated Dec 1, 2025 β’ 3.79k β’ 726
Running on Zero Agents Featured 455 DeepSeek OCR Demo π 455 An interactive demo for the DeepSeek-OCR model.