Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
falamarcao 's Collections
Evaluation
World Models
Medical
segmentation
Testing 1.. 2… 3…
AI Agent
Veterinary
On Device (local)
Start Here
omni models (text, image, audio, video)
Speech related
Web GPU
Software Engineering
Tracker
Speech-to-speech
MCP Servers
computer-use
Speech-to-text
Index-embed
3D
Code
Object Detection
Safety
Parser
Multimodal
Specialized
OCR
Video
Image
Audio
LLM
Text-to-speech

Speech-to-text

updated Jun 24
Upvote
-

  • nvidia/parakeet-tdt-0.6b-v2

    Automatic Speech Recognition • Updated Jun 29 • 642k • 1.53k

  • Running on Zero
    Agents
    Featured
    480

    Parakeet-TDT-0.6b-V2

     
    480

    Transcribe audio files with timestamps and downloadable subtitles


  • Runtime error
    Agents
    33

    Blazing Fast Whisper

    👁
    33

    Blazing Fast Whisper Deployed on HF Inference Endpoints


  • Running on CPU Upgrade
    Agents
    Featured
    1.42k

    Open ASR Leaderboard

    🏆
    1.42k

    Compare speech-to-text models across languages and datasets


  • LiquidAI/LFM2.5-Audio-1.5B

    Audio-to-Audio • 1B • Updated Mar 30 • 1.02k • 451

  • nvidia/nemotron-3.5-asr-streaming-0.6b

    Automatic Speech Recognition • 0.6B • Updated 4 days ago • 1.04M • • 1.01k
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs