StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents Paper β’ 2607.22798 β’ Published Jul 24 β’ 63
Shaping capabilities with token-level data filtering Paper β’ 2601.21571 β’ Published Jan 29 β’ 32
DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs Paper β’ 2601.03559 β’ Published Jan 7 β’ 14
Running on CPU Upgrade Featured 3.31k The Smol Training Playbook π 3.31k The secrets to building world-class LLMs
Video models are zero-shot learners and reasoners Paper β’ 2509.20328 β’ Published Sep 24, 2025 β’ 101
view article Article Introducing Trackio: A Lightweight Experiment Tracking Library from Hugging Face +3 abidlabs, znation, nouamanetazi, sasha, qgallouedec β’ Jul 29, 2025 β’ 228
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach Paper β’ 2502.05171 β’ Published Feb 7, 2025 β’ 162