OpenMOSS-Team/MOSS-VL-Instruct-0708-NF4 Video-Text-to-Text • 11B • Updated 1 day ago • 89 • 14
OpenMOSS-Team/MOSS-VL-Instruct-0708-FP8 Video-Text-to-Text • 11B • Updated 1 day ago • 95 • 18
OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation Paper • 2607.23855 • Published 24 days ago • 27
World Action Models: The Next Frontier in Embodied AI Paper • 2605.12090 • Published May 12 • 71
The Past Is Not Past: Memory-Enhanced Dynamic Reward Shaping Paper • 2604.11297 • Published Apr 13 • 144