facebook/wav2vec2-base-10k-voxpopuli-ft-fr
Automatic Speech Recognition • Updated • 50
None defined yet.
MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding
Skaling: Chinchilla's Exponents Meet Kaplan's Coupling