-
MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Paper • 2306.00107 • Published • 6 -
MusiLingo: Bridging Music and Text with Pre-trained Language Models for Music Captioning and Query Response
Paper • 2309.08730 • Published • 2 -
ChatMusician: Understanding and Generating Music Intrinsically with LLM
Paper • 2402.16153 • Published • 57 -
CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
Paper • 2401.11944 • Published • 27
Collections
Discover the best community collections!
Collections trending this week
-
NbAiLab/nb-whisper-large-distil-turbo-beta
Automatic Speech Recognition • 0.8B • Updated • 185 • 12 -
NbAiLab/nb-whisper-large
Automatic Speech Recognition • 2B • Updated • 4.52k • 46 -
NbAiLab/nb-whisper-medium
Automatic Speech Recognition • 0.8B • Updated • 827 • 4 -
NbAiLab/nb-whisper-small
Automatic Speech Recognition • 0.2B • Updated • 2.28k • 2
-
The Goldilocks of Pragmatic Understanding: Fine-Tuning Strategy Matters for Implicature Resolution by LLMs
Paper • 2210.14986 • Published • 5 -
Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
Paper • 2311.10702 • Published • 19 -
Large Language Models as Optimizers
Paper • 2309.03409 • Published • 79 -
From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting
Paper • 2309.04269 • Published • 34
-
nvidia/parakeet-rnnt-1.1b
Automatic Speech Recognition • 1B • Updated • 3.33k • 185 -
nvidia/parakeet-ctc-1.1b
Automatic Speech Recognition • 1B • Updated • 217k • 61 -
nvidia/parakeet-rnnt-0.6b
Automatic Speech Recognition • 0.6B • Updated • 56.2k • 15 -
nvidia/parakeet-ctc-0.6b
Automatic Speech Recognition • 0.6B • Updated • 62.4k • 27
-
MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Paper • 2306.00107 • Published • 6 -
MusiLingo: Bridging Music and Text with Pre-trained Language Models for Music Captioning and Query Response
Paper • 2309.08730 • Published • 2 -
ChatMusician: Understanding and Generating Music Intrinsically with LLM
Paper • 2402.16153 • Published • 57 -
CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
Paper • 2401.11944 • Published • 27
-
NbAiLab/nb-whisper-large-distil-turbo-beta
Automatic Speech Recognition • 0.8B • Updated • 185 • 12 -
NbAiLab/nb-whisper-large
Automatic Speech Recognition • 2B • Updated • 4.52k • 46 -
NbAiLab/nb-whisper-medium
Automatic Speech Recognition • 0.8B • Updated • 827 • 4 -
NbAiLab/nb-whisper-small
Automatic Speech Recognition • 0.2B • Updated • 2.28k • 2
-
nvidia/parakeet-rnnt-1.1b
Automatic Speech Recognition • 1B • Updated • 3.33k • 185 -
nvidia/parakeet-ctc-1.1b
Automatic Speech Recognition • 1B • Updated • 217k • 61 -
nvidia/parakeet-rnnt-0.6b
Automatic Speech Recognition • 0.6B • Updated • 56.2k • 15 -
nvidia/parakeet-ctc-0.6b
Automatic Speech Recognition • 0.6B • Updated • 62.4k • 27
-
The Goldilocks of Pragmatic Understanding: Fine-Tuning Strategy Matters for Implicature Resolution by LLMs
Paper • 2210.14986 • Published • 5 -
Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
Paper • 2311.10702 • Published • 19 -
Large Language Models as Optimizers
Paper • 2309.03409 • Published • 79 -
From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting
Paper • 2309.04269 • Published • 34