Large-scale Contrastive Language-Audio Pretraining with Feature Fusion and Keyword-to-Caption Augmentation Paper • 2211.06687 • Published Nov 12, 2022 • 7
Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published Jul 10 • 87
CohereLabs/cohere-transcribe-arabic-07-2026 Automatic Speech Recognition • 2B • Updated 28 days ago • 48.8k • 163
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published Jul 2 • 309
Towards Automating Scientific Review with Google's Paper Assistant Tool Paper • 2606.28277 • Published Jun 26 • 11