Running Featured 94 Distilling 100B+ Models 40x Faster with TRL 📝 94 TRL distillation for 100B+ teachers, 40x faster
Runtime error MCP Featured 127 Mage-Flow 🎨 127 Efficient native-resolution image generation and editing
Running 67 Don't Train the Model, Evolve the Harness 🌿 67 Evolving an agent's harness, not its model, on Harvey's LAB
Running 245 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 245 Building and scaling RL environments for LLM training
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 15.3k • • 2.94k
Running on CPU Upgrade 281 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 281 Visualize synthetic‑data experiments as an interactive bookshelf
mistralai/Voxtral-Mini-4B-Realtime-2602 Automatic Speech Recognition • 4B • Updated Mar 11 • 1.78M • 994
Running Featured 1.45k FineWeb: decanting the web for the finest text data at scale 🍷 1.45k Explore and download the FineWeb web‑scale text dataset
Running 4.05k The Ultra-Scale Playbook 🌌 4.05k The ultimate guide to training LLM on large GPU Clusters
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook 📚 3.31k The secrets to building world-class LLMs