Visual Question Answering
Transformers
Safetensors
English
videollama2_mistral
text-generation
multimodal large language model
large video-language model
Instructions to use Aliayub1995/VideoLLaMA2-7B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Aliayub1995/VideoLLaMA2-7B with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("visual-question-answering", model="Aliayub1995/VideoLLaMA2-7B")# Load model directly from transformers import AutoModelForCausalLM model = AutoModelForCausalLM.from_pretrained("Aliayub1995/VideoLLaMA2-7B", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update videollama2/mm_utils.py
Browse files- videollama2/mm_utils.py +1 -0
videollama2/mm_utils.py
CHANGED
|
@@ -5,6 +5,7 @@ import base64
|
|
| 5 |
import traceback
|
| 6 |
from io import BytesIO
|
| 7 |
import gdown
|
|
|
|
| 8 |
|
| 9 |
import cv2
|
| 10 |
import torch
|
|
|
|
| 5 |
import traceback
|
| 6 |
from io import BytesIO
|
| 7 |
import gdown
|
| 8 |
+
import logging
|
| 9 |
|
| 10 |
import cv2
|
| 11 |
import torch
|