— Journal Club 📚 — Collection Candidate papers to read in the Hugging Face journal club • 55 items • Updated 28 days ago • 39
Running 230 The ultimate guide to RL environments: building and scaling them in the LLM era 📝 230 Building and scaling RL environments for LLM training
view article Article Keep the Tokens Flowing: Lessons from 16 Open-Source RL Libraries +7 aminediroHF, qgallouedec, kashif, lewtun, edbeeching, albertvillanova, nouamanetazi, lvwerra, sergiopaniego • Mar 10 • 182
Running 119 Unlocking On-Policy Distillation for Any Model Family 📝 119 Explore on-policy distillation visualization for any model
Running Featured 93 Distilling 100B+ Models 40x Faster with TRL 📝 93 TRL distillation for 100B+ teachers, 40x faster
view article Article Welcome Gemma 4: Frontier multimodal intelligence on device +5 merve, pcuenq, sergiopaniego, burtenshaw, Steveeeeeeen, alvarobartt, SaylorTwift • Apr 2 • 925
Running on CPU Upgrade Agents 1.03k Open VLM Leaderboard 🌎 1.03k VLMEvalKit Evaluation Results Collection
view article Article Introducing smolagents: simple agents that write actions in code. +1 m-ric, merve, thomwolf • Dec 31, 2024 • 1.21k
view article Article Vision Language Models (Better, faster, stronger) +3 merve, sergiopaniego, ariG23498, pcuenq, andito • May 12, 2025 • 616
view article Article 视觉语言模型 (更好、更快、更强) +3 merve, sergiopaniego, ariG23498, pcuenq, andito • May 12, 2025 • 19
Running on CPU Upgrade 275 The Synthetic Data Playbook: Generating Trillions of the Finest Tokens 📝 275 Visualize synthetic‑data experiments as an interactive bookshelf
Running Featured 84 QED-Nano: Teaching a Tiny Model to Prove Hard Theorems 📝 84 Who needs 1T parameters? Olympiad proofs with a 4B model