view article Article 🌀 RoPE (Rotary Position Embedding) — When AI finally learns where it is! 📍✨ RDTvlokip • Sep 30, 2025 • 4
DFlash Collection Block Diffusion for Flash Speculative Decoding • 23 items • Updated 22 days ago • 156
Programming with Data: Test-Driven Data Engineering for Self-Improving LLMs from Raw Corpora Paper • 2604.24819 • Published Apr 27 • 91
view post Post 1753 Interested in RL training environments?We just released a beginner-friendly walkthrough notebook!Train a model to play Wordle using TRL + OpenEnv (TextArena) + GRPO + vLLM.happy learning! 🌱Notebook: https://github.com/huggingface/trl/blob/main/examples/notebooks/openenv_wordle_grpo.ipynbOpenEnv guide in TRL: https://huggingface.co/docs/trl/main/en/openenv See translation 👍 7 7 + Reply