Speech Collection by asdkazmi Dec 5, 2024 - facebook/seamless-m4t-v2-large Automatic Speech Recognition • 2B • Updated Jan 4, 2024 • 297k • 999 rishitdagli/see-2-sound Updated Jul 6, 2024 • 7 • 8 ai4bharat/indic-parler-tts Text-to-Speech • 0.9B • Updated Sep 24, 2025 • 285k • 280 parler-tts/parler-tts-large-v1 Text-to-Speech • 2B • Updated Nov 22, 2024 • 11.8k • 274
Visual LLMs Collection by sekosan Jan 25 - liuhaotian/llava-v1.5-13b-lora Image-Text-to-Text • Updated May 9, 2024 • 17 • 28 zai-org/RealVideo Any-to-Any • Updated Dec 11, 2025 • 105 FlowAct-R1: Towards Interactive Humanoid Video Generation Paper • 2601.10103 • Published Jan 15 • 78
LLM Agents Collection by Yankz Jan 4, 2024 - GPT-4V(ision) is a Generalist Web Agent, if Grounded Paper • 2401.01614 • Published Jan 3, 2024 • 23
Transcription Collection by pranavagrawal Jan 4, 2024 - Runtime error Agents 9 Vakyansh Hindi Speech Recognition Demo (Wav2vec2) 🐠 9
papers Collection by rodchile Jan 4, 2024 - LLM in a flash: Efficient Large Language Model Inference with Limited Memory Paper • 2312.11514 • Published Dec 12, 2023 • 264
LLM in a flash: Efficient Large Language Model Inference with Limited Memory Paper • 2312.11514 • Published Dec 12, 2023 • 264
AUTOM1111-EXTENSIONS Collection by Steveefemsc Nov 18, 2024 - kabachuha/modelscope-damo-text2video-pruned-weights Updated Mar 23, 2023 • 816 • 39 ali-vilab/modelscope-damo-text-to-video-synthesis Text-to-Video • Updated Mar 29, 2023 • 124 • 482 h94/IP-Adapter-FaceID Text-to-Image • Updated Apr 16, 2024 • 215k • 1.87k SWivid/F5-TTS Text-to-Speech • Updated Mar 21, 2025 • 770k • 1.19k
3D Modeling Collection by StrangeBoltz Jan 4, 2024 - Paint3D: Paint Anything 3D with Lighting-Less Texture Diffusion Models Paper • 2312.13913 • Published Dec 21, 2023 • 24
Paint3D: Paint Anything 3D with Lighting-Less Texture Diffusion Models Paper • 2312.13913 • Published Dec 21, 2023 • 24
sd Collection by aerosxue Jan 4, 2024 - Running on Zero MCP Featured 2.03k Stable Video Diffusion 1.1 📺 2.03k Generate a short video from a single image Runtime error Agents 49 Fast SDXL 🔥 49
Running on Zero MCP Featured 2.03k Stable Video Diffusion 1.1 📺 2.03k Generate a short video from a single image
Speech Collection by asdkazmi Dec 5, 2024 - facebook/seamless-m4t-v2-large Automatic Speech Recognition • 2B • Updated Jan 4, 2024 • 297k • 999 rishitdagli/see-2-sound Updated Jul 6, 2024 • 7 • 8 ai4bharat/indic-parler-tts Text-to-Speech • 0.9B • Updated Sep 24, 2025 • 285k • 280 parler-tts/parler-tts-large-v1 Text-to-Speech • 2B • Updated Nov 22, 2024 • 11.8k • 274
Visual LLMs Collection by sekosan Jan 25 - liuhaotian/llava-v1.5-13b-lora Image-Text-to-Text • Updated May 9, 2024 • 17 • 28 zai-org/RealVideo Any-to-Any • Updated Dec 11, 2025 • 105 FlowAct-R1: Towards Interactive Humanoid Video Generation Paper • 2601.10103 • Published Jan 15 • 78
papers Collection by rodchile Jan 4, 2024 - LLM in a flash: Efficient Large Language Model Inference with Limited Memory Paper • 2312.11514 • Published Dec 12, 2023 • 264
LLM in a flash: Efficient Large Language Model Inference with Limited Memory Paper • 2312.11514 • Published Dec 12, 2023 • 264
AUTOM1111-EXTENSIONS Collection by Steveefemsc Nov 18, 2024 - kabachuha/modelscope-damo-text2video-pruned-weights Updated Mar 23, 2023 • 816 • 39 ali-vilab/modelscope-damo-text-to-video-synthesis Text-to-Video • Updated Mar 29, 2023 • 124 • 482 h94/IP-Adapter-FaceID Text-to-Image • Updated Apr 16, 2024 • 215k • 1.87k SWivid/F5-TTS Text-to-Speech • Updated Mar 21, 2025 • 770k • 1.19k
LLM Agents Collection by Yankz Jan 4, 2024 - GPT-4V(ision) is a Generalist Web Agent, if Grounded Paper • 2401.01614 • Published Jan 3, 2024 • 23
3D Modeling Collection by StrangeBoltz Jan 4, 2024 - Paint3D: Paint Anything 3D with Lighting-Less Texture Diffusion Models Paper • 2312.13913 • Published Dec 21, 2023 • 24
Paint3D: Paint Anything 3D with Lighting-Less Texture Diffusion Models Paper • 2312.13913 • Published Dec 21, 2023 • 24
Transcription Collection by pranavagrawal Jan 4, 2024 - Runtime error Agents 9 Vakyansh Hindi Speech Recognition Demo (Wav2vec2) 🐠 9
sd Collection by aerosxue Jan 4, 2024 - Running on Zero MCP Featured 2.03k Stable Video Diffusion 1.1 📺 2.03k Generate a short video from a single image Runtime error Agents 49 Fast SDXL 🔥 49
Running on Zero MCP Featured 2.03k Stable Video Diffusion 1.1 📺 2.03k Generate a short video from a single image