Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
syddharth
's Collections
LLM
Audio
Video
Image
Vision
Vision
updated
Sep 23, 2024
Upvote
-
Sort: Collection
01-ai/Yi-VL-34B
Image-Text-to-Text
•
Updated
Jun 26, 2024
•
224
•
265
01-ai/Yi-VL-6B
Image-Text-to-Text
•
Updated
Jun 26, 2024
•
251
•
124
NousResearch/Nous-Hermes-2-Vision-Alpha
Text Generation
•
Updated
Dec 3, 2023
•
234
•
305
liuhaotian/llava-v1.5-13b
Image-Text-to-Text
•
Updated
May 9, 2024
•
16.1k
•
529
fancyfeast/joytag
Image Classification
•
91.5M
•
Updated
Mar 9, 2024
•
945
•
118
internlm/internlm-xcomposer2-7b
Text Generation
•
Updated
Feb 27, 2024
•
3.78k
•
31
internlm/internlm-xcomposer2-4khd-7b
Visual Question Answering
•
Updated
Apr 18, 2024
•
1.16k
•
73
SmilingWolf/wd-vit-large-tagger-v3
0.3B
•
Updated
Jul 26, 2024
•
864
•
92
Aryn/deformable-detr-DocLayNet
Object Detection
•
41.1M
•
Updated
Aug 8, 2025
•
26.7k
•
51
abetlen/Phi-3.5-vision-instruct-gguf
4B
•
Updated
Oct 1, 2024
•
1.07k
•
31
MiaoshouAI/Florence-2-base-PromptGen-v1.5
0.3B
•
Updated
Oct 9, 2024
•
1.08k
•
105
stepfun-ai/GOT-OCR2_0
Image-Text-to-Text
•
0.7B
•
Updated
Feb 4, 2025
•
554k
•
1.55k
Upvote
-
Sort: Collection
Share collection
View history
Collection guide
Browse collections