moondream2
a tiny vision language model
a tiny vision language model
MoonDream 2 Vision Model on the Browser: Candle/Rust/WASM
Generate text from images and prompts
Meta Llama3 8b with Llava Multimodal capabilities
Generate text and segment images using PaliGemma
Chat with an AI about any uploaded image
Chat with an image using Phi-3 Vision model
Microsoft Phi-3 Vision 128k with Multimodal capabilities
let's talk about the meaning of life
Convert images to grayscale
Generate captions, detections, and segmentations from images
Generate detailed captions for your images
Generate detailed captions for any image
Analyze images to detect objects, generate captions, or perform OCR
A private and powerful multimodal AI chatbot that runs local
Generate images from captions or enhanced prompts
Generate text based on an image and prompt
Ask questions about images
Ask questions about images and get detailed answers
GOT - OCR (from : UCAS, Beijing)
Chat with Llama about images and text
Huggingface space for JanusFlow-1.3B
Answer questions about images with AI chat
Generate captions for images