Hoi! - A Multimodal Dataset for Force-Grounded, Cross-View Articulated Manipulation Paper • 2512.04884 • Published Apr 15 • 3
Visual Representation Alignment for Multimodal Large Language Models Paper • 2509.07979 • Published Sep 9, 2025 • 84