Running 213 The ultimate guide to multi-harness RL 🔀 213 Train open models with RL inside real agent harnesses
Running 259 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 259 Building and scaling RL environments for LLM training
Running on CPU Upgrade 282 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 282 Explore synthetic data benchmarks with an interactive bookshelf
Running on CPU Upgrade Featured 3.32k The Smol Training Playbook 📚 3.32k The secrets to building world-class LLMs
Running on CPU Upgrade Agents Featured 1.49k Open ASR Leaderboard 🏆 1.49k Compare speech‑to‑text models across datasets