OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 22 days ago • 151
Running on Zero Agents 1.32k MiniMax H3 Turbo LoRA 🎬 1.32k Video generation with a synchronized soundtrack
GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 19 days ago • 132
Running 161 HF Viewer · Model Architecture Explorer 🟩 161 Interactive architecture graph for any HF model
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 30 days ago • 656
Running on Zero MCP Featured 2.77k Qwen-Image-Edit-2511-LoRAs-Fast 🎃 2.77k Demo of the Collection of Qwen Image Edit LoRAs
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published Jul 16 • 147
Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget Paper • 2607.13125 • Published Jul 18 • 140