Automatic Speech Recognition
Transformers
PyTorch
TensorFlow
JAX
Safetensors
whisper
audio
hf-asr-leaderboard
Eval Results (legacy)
Eval Results
Instructions to use openai/whisper-large with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use openai/whisper-large with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="openai/whisper-large")# Load model directly from transformers import AutoProcessor, AutoModelForSpeechSeq2Seq processor = AutoProcessor.from_pretrained("openai/whisper-large") model = AutoModelForSpeechSeq2Seq.from_pretrained("openai/whisper-large", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update config for automatic language detection
Browse filesThis should allow the model to detect the language by running a forward loop and filling with the `<lan>` lang token
- config.json +4 -0
config.json
CHANGED
|
@@ -30,6 +30,10 @@
|
|
| 30 |
],
|
| 31 |
[
|
| 32 |
2,
|
|
|
|
|
|
|
|
|
|
|
|
|
| 33 |
50363
|
| 34 |
]
|
| 35 |
],
|
|
|
|
| 30 |
],
|
| 31 |
[
|
| 32 |
2,
|
| 33 |
+
None
|
| 34 |
+
],
|
| 35 |
+
[
|
| 36 |
+
3,
|
| 37 |
50363
|
| 38 |
]
|
| 39 |
],
|