Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Jiaqi Song's picture

Jiaqi Song

Jiaqi1Song
5 2
·

AI & ML interests

Large Audio Language Model, Omni Model, NLP, World Model

Recent Activity

liked a dataset about 1 month ago
scbz/minspeech
upvoted a collection 2 months ago
NIM4-ASR
upvoted a paper 2 months ago
Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation
View all activity

Organizations

None yet

upvoted a collection 2 months ago

NIM4-ASR

Collection
NIM4-ASR is our in-house ASR model, optimized for parameter efficiency, hallucination mitigation, customization, and real-time streaming inference. • 1 item • Updated Jun 18 • 3
upvoted a paper 2 months ago

Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation

Paper • 2605.19833 • Published May 19 • 137
upvoted 2 collections 2 months ago

WenetSpeech-Wu

Collection
4 items • Updated Jan 31 • 6

Rethinking Entropy Allocation in LLM-based ASR

Collection
1 item • Updated Jun 11 • 2
upvoted a paper 2 months ago

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR

Paper • 2604.18105 • Published Apr 20 • 3
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs