Rethinking OPD Collection This collection includes the models used in the paper "Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recip • 5 items • Updated Jun 3 • 4
MaxRL Collection Qwen3-Base post-trained checkpoints for our paper, Maximum Likelihood Reinforcement Learning [https://zanette-labs.github.io/MaxRL/] • 5 items • Updated 20 days ago • 3
MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction Paper • 2606.18558 • Published Jun 17 • 53
VFIG Collection VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models • 3 items • Updated May 6 • 3