ATH-MaaS/Ovis-VL-Embedding-9B
Feature Extraction • Updated • 266 • 35
High-fidelity Text-To-Speech
Transcribe or translate audio and YouTube videos to text
Generate high‑resolution images from your text prompt
Extract text from images using OCR
Generate high‑resolution images from text prompts
Generate HTML code from a website screenshot
Generate personalized photos of a person from a prompt
Generate personalized images preserving your face identity
Generate a talking face video from an image and audio