Edit Models filters
Model Tree
Apps
Inference Providers
One-click Deployment
Models
3,397
Active filters: speech
didiudom94/whisper-small-ko-to-en-translator
Translation • Updated • 2
Louis0324/StyleStream
Updated • 1
Banaxi-Tech/BananaMind-TTS-V2
Text-to-Speech • 9.5M • Updated • 43 • 4
mlx-community/Irodori-TTS-600M-v3-VoiceDesign-fp16
Text-to-Speech • 0.6B • Updated • 68 • 1
mlx-community/Irodori-TTS-600M-v3-VoiceDesign-8bit
Text-to-Speech • 0.3B • Updated • 88 • 1
didiudom94/whisper-small-ko-to-en-v2-cross-attention
Translation • Updated • 2
faeea/custom-gopt-252-eval
Updated • 1
CogniSoftOrg/canary-1b-v2-mlx-bf16
Automatic Speech Recognition • 1.0B • Updated • 252 • 2
SpeechAntiSpoofingBenchmarks/ResCapsGuard
Updated • 3
SpeechAntiSpoofingBenchmarks/AASIST
Updated • 46 • 1
IHP-Lab/AF3_PCLM_DPO
Audio-Text-to-Text • Updated • 17 • 1
IHP-Lab/Qwen2-Audio_PCLM_DPO
Audio-Text-to-Text • 8B • Updated • 3.19k • 2
openbmb/MiniCPM-o-2_6-GPTQ
Any-to-Any • 9B • Updated • 79 • 4
Banaxi-Tech/BananaMind-TTS-V2.1-Preview
Text-to-Speech • Updated • 52 • 2
BUT-FIT/Dixtral_QA
Automatic Speech Recognition • 5B • Updated • 13
BUT-FIT/Dixtral
Automatic Speech Recognition • 5B • Updated • 49
EYEDOL/parakeet-tdt-0.6b-yoruba
Updated • 6
SpeechAntiSpoofingBenchmarks/AASIST-L
Updated • 12 • 1
SpeechAntiSpoofingBenchmarks/RawTFNet
Updated • 4
laion/vocalburst-locator
Audio Classification • Updated • 58 • 1
onnx-community/nemotron-3.5-asr-streaming-0.6b-onnx-int4
Automatic Speech Recognition • Updated • 1.75k • 18
SpeechAntiSpoofingBenchmarks/WhisperMFCCMesoNet
Updated • 24 • 1
omi-health/omi-med-stt-v1-mlx
Automatic Speech Recognition • 0.6B • Updated • 101 • 1
MIT-SLS/USAD2-Small
Feature Extraction • 24.8M • Updated • 74
SpeechAntiSpoofingBenchmarks/W2V2-AASIST
Updated • 6
MIT-SLS/USAD2-Base
Feature Extraction • 97.2M • Updated • 400
MIT-SLS/USAD2-Large
Feature Extraction • 0.3B • Updated • 452 • 1
MIT-SLS/USAD2-XLarge
Feature Extraction • 0.7B • Updated • 35 • 1
MIT-SLS/USAD2-Large-Plus
Feature Extraction • 0.3B • Updated • 337 • 1