view article Article autotrust/JEV-27B-VL: a decision model that learned to see without a single image of training autotrust • 11 days ago • 126
LEGO-Anything: Coding Agents for 3D Scene Reconstruction Paper • 2609.36380 • Published 14 days ago • 130
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 14 days ago • 322
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published Sep 10 • 657
SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem Paper • 2609.07064 • Published Sep 7 • 124