Video-Text-to-Text
Transformers
Safetensors
English
videollama3_qwen2
text-generation
multi-modal
large-language-model
video-language-model
custom_code
Instructions to use DAMO-NLP-SG/VideoLLaMA3-7B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use DAMO-NLP-SG/VideoLLaMA3-7B with Transformers:
# Load model directly from transformers import AutoModelForCausalLM model = AutoModelForCausalLM.from_pretrained("DAMO-NLP-SG/VideoLLaMA3-7B", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -40,8 +40,8 @@ base_model:
|
|
| 40 |
|
| 41 |
|
| 42 |
## π° News
|
| 43 |
-
<!-- * **[2024.01.23]** ππ Update technical report. If you have works closely related to VideoLLaMA3 but not mentioned in the paper, feel free to let us know.
|
| 44 |
-
* **[2024.01.
|
| 45 |
* **[2024.01.22]** Release models and inference code of VideoLLaMA 3.
|
| 46 |
|
| 47 |
## π Introduction
|
|
|
|
| 40 |
|
| 41 |
|
| 42 |
## π° News
|
| 43 |
+
<!-- * **[2024.01.23]** ππ Update technical report. If you have works closely related to VideoLLaMA3 but not mentioned in the paper, feel free to let us know. -->
|
| 44 |
+
* **[2024.01.24]** π₯π₯ Online Demo is available: [VideoLLaMA3-Image-7B](https://huggingface.co/spaces/lixin4ever/VideoLLaMA3-Image), [VideoLLaMA3-7B](https://huggingface.co/spaces/lixin4ever/VideoLLaMA3).
|
| 45 |
* **[2024.01.22]** Release models and inference code of VideoLLaMA 3.
|
| 46 |
|
| 47 |
## π Introduction
|