vid analysis
updated
prithivMLmods/Qwen3-VL-8B-Abliterated-Caption-it
Image-Text-to-Text
• 9B • Updated • 35
mradermacher/Qwen3-VL-8B-NSFW-Caption-V4.5-GGUF
8B • Updated • 12.1k
• 76
prithivMLmods/Qwen3-VisionCaption-2B
Image-Text-to-Text
• 2B • Updated • 19
• 5
msrcam/Qwen3-VL-2B-Instruct-heretic
ghost-actual/Qwen3.5-4B-Claude-Opus-4.6-Distilled-heretic
Text Generation
• 5B • Updated • 36
• 3
Image-Text-to-Text
• 3B • Updated • 28
• 5
Image-Text-to-Text
• 5B • Updated • 6.65M
• • 810
bobber/routangseng-qwen35-0.8b-abliterated-onnx
Image-Text-to-Text
• Updated • 21
bobber/routangseng-0.8b-hottake-onnx
Image-Text-to-Text
• Updated • 4
Caplin43/multimodal-vision-language-mini
Image-to-Text
• Updated • 17
Ytgetahun/visual-narrator-llm
Image-to-Text
• Updated
Ytgetahun/visual-narrator-vlm
0.2B • Updated • 3
allenai/MolmoPoint-Vid-4B
Video-Text-to-Text
• 5B • Updated • 898
• 13
TencentARC/ARC-Qwen-Video-7B-Narrator
Video-Text-to-Text
• 9B • Updated • 46
• 11
Video-Text-to-Text
• Updated • 80
• 32
Image-Text-to-Text
• 5B • Updated • 47.3k
• 52
DAMO-NLP-SG/VideoLLaMA3-2B
Video-Text-to-Text
• 2B • Updated • 3.01k
• 21
u94fmn391j/SAVANT-scene-description-lora
Image-to-Text
• Updated • 6
VINAY-UMRETHE/SigMamba-V1-Large
Video Classification
• 0.9B • Updated • 7
• 6
qoranet/QORA-Vision-Video
Video Classification
• Updated • 6
sumit7488/TimesFormer_Baseline
Video Classification
• 0.1B • Updated • 3
StreamFormer/streamformer-timesformer
Video Classification
• 0.1B • Updated • 110
• 4
facebook/vjepa2-vitg-fpc64-256
Video Classification
• 1B • Updated • 151k
• 57
Image-Text-to-Text
• 3B • Updated • 7.75k
• 201
Video-Text-to-Text
• Updated • 2
BidirLM/BidirLM-Omni-2.5B-Embedding
Sentence Similarity
• 2B • Updated • 456
• 45
Image-Text-to-Text
• Updated • 14
• 14
Video-Text-to-Text
• 0.9B • Updated • 8
• 1
Video-Text-to-Text
• 2B • Updated • 6.32k
• 580
LongVie 2: Multimodal Controllable Ultra-Long Video World Model
Paper
• 2512.13604
• Published • 76
Image-Text-to-Text
• 4B • Updated • 385k
• 2.9k
prithivMLmods/Gemma4-BLIP3o-Captioner-5B
Image-Text-to-Text
• 5B • Updated • 479
• 3
lewiswatson/Frame2KG-LFM-2.5-450m-JSON
Image-Text-to-Text
• 0.4B • Updated • 52
69.7M • Updated • 29.3k
• 2