MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training Paper • 2512.15411 • Published Dec 17, 2025
InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization Paper • 2607.04988 • Published Jul 6 • 28
WSA$_1$: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control Paper • 2607.03941 • Published Jul 4 • 1