view article Article What We Learned by Reproducing 2,200 papers from ICML abidlabs • 3 days ago • 62
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • 20 days ago • 469
Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks Paper • 2606.29082 • Published Jun 27 • 42
Running 62 Don't Train the Model, Evolve the Harness 🌿 62 Evolving an agent's harness, not its model, on Harvey's LAB
view article Article Profiling in PyTorch (Part 1): A Beginner's Guide to torch.profiler +3 ariG23498, sayakpaul, sergiopaniego, ror, pcuenq • May 29 • 159
Running 223 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 223 Building and scaling RL environments for LLM training
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe Paper • 2604.13016 • Published Apr 14 • 114
view article Article Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers tomaarsen • Apr 16 • 79
Running Featured 92 Distilling 100B+ Models 40x Faster with TRL 📝 92 TRL distillation for 100B+ teachers, 40x faster
view article Article TRL v1.0: Post-Training Library Built to Move with the Field +2 qgallouedec, stevhliu, pcuenq, sergiopaniego • Mar 31 • 58
view article Article Keep the Tokens Flowing: Lessons from 16 Open-Source RL Libraries +7 aminediroHF, qgallouedec, kashif, lewtun, edbeeching, albertvillanova, nouamanetazi, lvwerra, sergiopaniego • Mar 10 • 175
Running on CPU Upgrade 270 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 270 Visualize synthetic‑data experiments as an interactive bookshelf