Model Card for RL2-VLA QAM

For more information on usage, please refer to the RL2-VLA Github repository here.

Citation

@article{tan2026rl2,
  title   = {RL$^2$-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models},
  author  = {Derek Ming Siang Tan and Shailesh Shailesh and Srikrishna Iyer and William Wei Jie Teo and Yuanliang Ju and Qiao Gu and Guillaume Sartoretti},
  year    = {2026},
  journal = {arXiv preprint arXiv:2607.26991}
}
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Paper for rl2-vla/rl2-vla-qam-bridge