Video-Text-to-Text
Transformers
PyTorch
English
pixelrefer_qwen2
text-generation
multimodal large language model
large video-language model
Instructions to use Alibaba-DAMO-Academy/PixelRefer-2B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Alibaba-DAMO-Academy/PixelRefer-2B with Transformers:
# Load model directly from transformers import AutoModelForCausalLM model = AutoModelForCausalLM.from_pretrained("Alibaba-DAMO-Academy/PixelRefer-2B", dtype="auto") - Notebooks
- Google Colab
- Kaggle
Ctrl+K