music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 3.48k • 67 Running on Zero MCP 36 BS-Roformer Leap Audio Separator 🎵 36 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 18 StuPASE Speech Enhancement 🎙 18 Studio-quality generative speech enhancement
Running on Zero MCP 36 BS-Roformer Leap Audio Separator 🎵 36 Separate audio into vocals and instruments with BS-Roformer
Image Qwen/Qwen-Image-Edit Image-to-Image • 20B • Updated Aug 25, 2025 • 113k • • 2.53k sensenova/SenseNova-U1.5-8B-MoT Any-to-Any • 18B • Updated 13 days ago • 7.02k • 260 Qwen/Qwen-Image-2.1 Text-to-Image • 7B • Updated 7 days ago • 58.7k • 2.54k
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 323 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 54
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 54
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 848k • 1.11k zai-org/GLM-OCR Image-to-Text • 1B • Updated 17 days ago • 1.79M • 2.09k uv-scripts/ocr Updated 4 days ago • 4.59k • 163 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 414k • 496
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 2.51k • 744 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 149 • 610 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 61.5k • • 789 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated Aug 14 • 6.84M • • 16.5k
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 717k • 2.49k Running Featured 450 FastVLM WebGPU 🍎 450 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated Aug 18 • 4.54k • 821 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Model training merve/smol-vision Image-Text-to-Text • Updated Aug 10 • 196 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 233 • 217 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 15.3k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 968
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 15.3k • • 2.94k
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 3.48k • 67 Running on Zero MCP 36 BS-Roformer Leap Audio Separator 🎵 36 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 18 StuPASE Speech Enhancement 🎙 18 Studio-quality generative speech enhancement
Running on Zero MCP 36 BS-Roformer Leap Audio Separator 🎵 36 Separate audio into vocals and instruments with BS-Roformer
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 848k • 1.11k zai-org/GLM-OCR Image-to-Text • 1B • Updated 17 days ago • 1.79M • 2.09k uv-scripts/ocr Updated 4 days ago • 4.59k • 163 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 414k • 496
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 2.51k • 744 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 149 • 610 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 61.5k • • 789 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated Aug 14 • 6.84M • • 16.5k
Image Qwen/Qwen-Image-Edit Image-to-Image • 20B • Updated Aug 25, 2025 • 113k • • 2.53k sensenova/SenseNova-U1.5-8B-MoT Any-to-Any • 18B • Updated 13 days ago • 7.02k • 260 Qwen/Qwen-Image-2.1 Text-to-Image • 7B • Updated 7 days ago • 58.7k • 2.54k
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 717k • 2.49k Running Featured 450 FastVLM WebGPU 🍎 450 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated Aug 18 • 4.54k • 821 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 323 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 54
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 54
Model training merve/smol-vision Image-Text-to-Text • Updated Aug 10 • 196 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 233 • 217 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 15.3k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 968
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 15.3k • • 2.94k