Voice Activity Detection Collection Speech/non-speech segmentation models: ASR end-pointing, diarization pre-processing, streaming turn detection. Server-side and on-device/browser. • 8 items • Updated 15 days ago
Running on Zero MCP Featured 133 SenseNova-U1.5-8B-MoT 🎨 133 Unified text-to-image and image editing model
openai/whisper-large-v3-turbo Automatic Speech Recognition • 0.8B • Updated Oct 4, 2024 • 6.7M • • 3.37k