Video-Text-to-Text
Transformers
Safetensors
English
Chinese
moss_vl
feature-extraction
MOSS-VL
image-understanding
video-understanding
FP8
compressed-tensors
quantized
SGLang
custom_code
Instructions to use OpenMOSS-Team/MOSS-VL-Instruct-0708-FP8 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use OpenMOSS-Team/MOSS-VL-Instruct-0708-FP8 with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("OpenMOSS-Team/MOSS-VL-Instruct-0708-FP8", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
| { | |
| "_from_model_config": true, | |
| "bos_token_id": 151643, | |
| "eos_token_id": 151645, | |
| "transformers_version": "4.57.1", | |
| "cache_implementation": "quantized", | |
| "cache_config": { | |
| "backend": "hqq", | |
| "nbits": 8, | |
| "axis_key": 0, | |
| "axis_value": 0, | |
| "q_group_size": 64, | |
| "residual_length": 128 | |
| } | |
| } | |