Automatic Speech Recognition
Transformers
Safetensors
English
Chinese
voxtral_realtime
asr
speech-recognition
streaming-asr
speech
audio
multimodal
voxtral
chinese
english
code-switching
Instructions to use x-square-robot/X2Streaming-ASR-4B-1009 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use x-square-robot/X2Streaming-ASR-4B-1009 with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="x-square-robot/X2Streaming-ASR-4B-1009")# pip install -U transformers accelerate # Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("x-square-robot/X2Streaming-ASR-4B-1009") model = AutoModelForMultimodalLM.from_pretrained("x-square-robot/X2Streaming-ASR-4B-1009", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Ctrl+K