SenseVoiceSmall FunAudioLLM/SenseVoiceSmall Automatic Speech Recognition • Updated Jun 20 • 17.7k • 492
SmolVLM-Instruct HuggingFaceTB/SmolVLM-500M-Instruct Image-Text-to-Text • 0.5B • Updated Apr 8, 2025 • 133k • 199
Qwen2.5-VL-3B-Instruct Qwen/Qwen2.5-VL-3B-Instruct Image-Text-to-Text • 4B • Updated Apr 6, 2025 • 2.25M • • 711
SenseVoiceSmall FunAudioLLM/SenseVoiceSmall Automatic Speech Recognition • Updated Jun 20 • 17.7k • 492
SmolVLM-Instruct HuggingFaceTB/SmolVLM-500M-Instruct Image-Text-to-Text • 0.5B • Updated Apr 8, 2025 • 133k • 199
Qwen2.5-VL-3B-Instruct Qwen/Qwen2.5-VL-3B-Instruct Image-Text-to-Text • 4B • Updated Apr 6, 2025 • 2.25M • • 711