Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model Paper • 2609.18323 • Published 23 days ago • 133
Where Success Breaks: Failure-Boundary Learning for Robust Vision-Language-Action Models Paper • 2609.06114 • Published Sep 5
Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance Paper • 2603.02175 • Published Mar 2 • 24
UniCode$^2$: Cascaded Large-scale Codebooks for Unified Multimodal Understanding and Generation Paper • 2506.20214 • Published Jun 25, 2025 • 2
Code2Video: A Code-centric Paradigm for Educational Video Generation Paper • 2510.01174 • Published Oct 1, 2025 • 36
RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning Paper • 2504.18904 • Published Apr 26, 2025 • 10