Running on CPU Upgrade Agents 400 Deep Reinforcement Learning Leaderboard π 400 Search your models on the Deep RL leaderboard
HuggingFaceTB/SmolLM2-135M-Instruct Text Generation β’ 0.1B β’ Updated Sep 22, 2025 β’ 1.83M β’ 428