Chuntao Dan
p051tr0n
·
AI & ML interests
all kinds
Organizations
Vision
-
Zigeng/SlimSAM-uniform-50
Mask Generation • 28M • Updated • 1.72k • 15 -
facebook/detr-resnet-50
Object Detection • 41.6M • Updated • 349k • • 977 -
Intel/dpt-hybrid-midas
Depth Estimation • Updated • 717k • 110 -
openai/clip-vit-large-patch14
Zero-Shot Image Classification • 0.4B • Updated • 8.65M • 2.1k
Robot
Voice
Multimodal
-
Salesforce/blip-itm-base-coco
Updated • 19.5k • 29 -
Salesforce/blip-image-captioning-base
Image-to-Text • Updated • 1.66M • 892 -
Salesforce/blip-vqa-base
Visual Question Answering • 0.4B • Updated • 608k • 195 -
openai/clip-vit-large-patch14
Zero-Shot Image Classification • 0.4B • Updated • 8.65M • 2.1k
Agentic
Voice
Vision
-
Zigeng/SlimSAM-uniform-50
Mask Generation • 28M • Updated • 1.72k • 15 -
facebook/detr-resnet-50
Object Detection • 41.6M • Updated • 349k • • 977 -
Intel/dpt-hybrid-midas
Depth Estimation • Updated • 717k • 110 -
openai/clip-vit-large-patch14
Zero-Shot Image Classification • 0.4B • Updated • 8.65M • 2.1k
Multimodal
-
Salesforce/blip-itm-base-coco
Updated • 19.5k • 29 -
Salesforce/blip-image-captioning-base
Image-to-Text • Updated • 1.66M • 892 -
Salesforce/blip-vqa-base
Visual Question Answering • 0.4B • Updated • 608k • 195 -
openai/clip-vit-large-patch14
Zero-Shot Image Classification • 0.4B • Updated • 8.65M • 2.1k
Robot