Ai
-
Text-to-Speech β’ Updated β’ 180 β’ 789 -
LoRA Studio
πͺ618Browse and run thousands of community trained LoRAs
-
meta-llama/Llama-2-70b-hf
Text Generation β’ 69B β’ Updated β’ 4.87k β’ 854 -
meta-llama/Llama-2-13b-chat
Text Generation β’ Updated β’ 9 β’ 295 -
Whisper Speech X DreamTalk
π½179Combine voice cloning and portrait lipsync animation
-
Canary 1b
π€197Transcribe and translate audio into text
-
OutfitAnyone
π’2.81kGenerate virtual tryβon images for any model and clothing
-
YOLO World
π₯490Detect objects in images or videos
-
Face Liveness Detection SDK
π116FaceOnLive On-Premise Solution
-
ID Document Recognition SDK
πͺͺ346FaceOnLive On-Premise Solution
-
Arena Leaderboard
π4.97kView the LMArena leaderboard in fullβscreen
-
AI Comic Factory
π©11.2kCreate your own AI comic with a single prompt
-
MLLM-guided Image Editing (MGIE)
π©332Edit images using naturalβlanguage instructions
-
Qwen/Qwen-VL-Chat
Text Generation β’ Updated β’ 21.4k β’ 384 -
NIST FRVT TOP 1 Face Recognition, Face Liveness Detection, Face Analysis
π₯600Compare two faces to verify identity
-
Stable Cascade
π1.68kGenerate highβresolution images from text prompts
-
Remove Background Web
πΌ760In-browser background removal
-
unity/inference-engine-jets-text-to-speech
Updated β’ 63 β’ 16 -
Transformers.js
π1Upload an image to detect objects
-
SALMONN Audio Questioning
β‘85Deeply interrogate audio file content
-
TransferAnything
π’345Create custom images by transferring layout, color, and style
-
EVACLIP
π53Comparing powerful zero-shot image classification models
-
YOLO-World + EfficientSAM
π₯259Detect and segment objects in images or videos
-
PhotoMaker
π·1.96kGenerate personalized photos of a person from a prompt
-
PhotoMaker Style
π·655Generate personalized stylized portraits from your photos
-
LGM
π¦327Generate 3D models and videos from text or images
-
Adaptive Retrieval Web
π₯75Retrieve relevant answers with transformer-powered search
-
MusicGen
π΅5.08kGenerate music from a text description and optional melody
-
TTS Arena V2
π£970Compare and rank TTS voices by listening and voting
-
Stable Video Diffusion 1.1
πΊ2.03kGenerate a short video from a single image
-
Supa Fast Image Variations
β‘79Explore image variations from a single conceptual image
-
XTTS
πΈ2.77kGenerate speech from text using a reference voice
-
BlogWriterApp
π37Generate a blog post with AI
-
PDF Chatbot
π379Ask questions about PDFs using a chatbot
-
Video Face Swap
π±1.46kVideo deep fake
-
Gligen Demo
π174 -
Whisper JAX
β‘2.84k -
Mixtral-46.7B
π¦439Generate text responses to your queries
-
OpenVoice
π€1.14kGenerate speech in a cloned voice from a short audio clip
-
MetaVoice 1B
π£144A demo of MetaVoice 1B, a new TTS model by MetaVoice.
-
Zephyr Chat
πͺ904Chat with an AI model
-
LoRA the Explorer SDXL
π1.17kExplore fun LoRAs and generate with SDXL
-
Owl Tracking
β‘64Powerful foundation model for zero-shot object tracking
-
AnimateLCM SVD
π’356Generate animated video from a single image
-
Playground V2.5
π1.14kGenerate highly aesthetic images
-
ToDo
β‘21 -
Voice Cloning
β‘545Clone a voice and generate speech
-
Samba CoE V0.1
π39Generate responses in a chat format
-
Image Upscaling Playground
π¦855Upscale images with selectable AI models
-
fka/prompts.chat
Viewer β’ Updated β’ 2.1k β’ 34k β’ 9.78k -
Laiyer Deberta V3 Base Prompt Injection
π3Answer questions with context
-
ChatGPT Prompt Generator
π¨14 -
bigcode/starcoder2-15b
Text Generation β’ 16B β’ Updated β’ 5.62k β’ 674 -
InferenceClient Chatbots
π10 -
openai/whisper-large-v3
Automatic Speech Recognition β’ 2B β’ Updated β’ 5.1M β’ β’ 6.12k -
Whisper Large V3
π€«850Transcribe audio or YouTube videos to text
-
MagicAnimate
π1.43kGenerate animated videos from images and motion sequences
-
Ask AI over Youtube video
π½81Ask questions about YouTube videos
-
MeloTTS
π£478Fast, efficient, & multilingual text-to-speech
-
StarCoder2
β¨80Latest Coding Model Byπ€Huggingface & Friends-101 languages!
-
Bark
πΆ2.38kGenerate realistic speech and sounds from typed text
-
Face Swap
π©809Image deep fake (uncensored)
-
Animagine XL 3.1
π1.41kThe most opinionated, anime-themed SDXL model
-
VideoMamba
π97Identify actions and objects in videos and images
-
ComfyUI Laucher
πͺ154Launch the ComfyUI interface from your browser
-
VisualStylePrompting
π55 -
Whisper
π2.82kTranscribe audio or YouTube video into text
-
NaturalSpeech3 FACodec
π177Reconstruct speech and change voice style
-
DTG Demo
π103Generate refined Danbooru tags for image prompts
-
Audio Editing
π§329Edit audios with text prompts
-
Swap Face Model
π»240Swap faces in photos with enhanced results
-
Musiclang
β‘149 -
LD T3D
π³42 -
CodeFormer
πΌ2.4kRestore and enhance faces in photos with optional upscaling
-
GRM
π85Display a live demo website
-
SDXL-Lightning
β‘312Generate fast images from text prompts
-
GeoWizard
π¨125Generate depth, normals, and 3D model from a single image
-
QR Code AI Art Generator
π±1.99kQR Code AI Art Generator Blend QR codes with AI Art
-
ConsistI2V
π₯35Image to Video Synthesis
-
Image-Prompter
π98 -
Can You Run It? LLM version
π1.05kCheck if your GPU can run a chosen LLM model
-
Depth Anything
π566Generate depth map from a single image
-
RVC Inference HF
π485Combine and process audio files with effects
-
Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference
Paper β’ 2403.14520 β’ Published β’ 35 -
StyleTTS 2
π£733Efficient, fast, and natural text to speech with StyleTTS 2!
-
4DGS Demo
π115View interactive 3D scenes directly in your browser
-
Ratchet + Whisper
π£70Convert audio to text
-
Arc2Face
π₯168Generate new images of the same person from a face photo
-
DALLE 3 XL v2
π₯1 -
DALLE 3 XL v2
π₯1.79kGenerate highβresolution images from your text prompt
-
llava-hf/vip-llava-13b-hf
Image-Text-to-Text β’ 13B β’ Updated β’ 40 β’ 12 -
text-generation-webui
π9 -
BrushNet
β‘176 -
SDXS-512-0.9 GPU Demo -1 Steps
β‘84SDXS 1024 will Release this is based on SDXS-512-0.9
-
DALLΒ·E mini
π₯5.71kGenerate images from any text prompt
-
ASR High Accuracy Test
π’2 -
Seamless M4T v2
π517Translate speech and text between languages
-
Multitrack Midi Music Generator
π΅125Generate multitrack music with selectable genre and tempo
-
Screenshot to HTML
β‘928Generate HTML code from a website screenshot
-
Idefics 8b
π146Generate text from images and prompts
-
Chat With Llama3 8b
π399Latest text-generation model by META - Meta Llama3 8b.
-
β AI Jukebox β
πΆ414Generate music powered by AI
-
β Hub API Playground β
πΉ138Try the Hugging Face API through the playground
-
T2I-Adapter-SDXL
π290Generate images from text prompts
-
Stable Diffusion Web UI
π§591Set up and customize Stable Diffusion WebUI
-
Recommend Similar Papers
π195Generate similar paper recommendations from a Hugging Face link
-
Invisible Stitch
πͺ‘45Generate 3D scenes from images with prompts
-
InstantStyle
π458Style-Preserving Text-to-Image Generation
-
Hyper SD15 Scribble
π₯198Generate images from sketches and text prompts
-
Ugen Image Captioning
π48 -
PixArt Sigma 900M
πΌ153Generate images from text prompts
-
Live Portrait
π€ͺ3.78kApply the motion of a video on a portrait
-
Llama3.1 405B
π776Generate responses with Llamaβ―3.1β―405B AI model
-
Video Dubbing (SoniTranslate)
π854Video Dubbing with Open Source Projects
-
Whisper-all-zero
π49Transcribe audio into text using different models
-
Turbo Edit
π132Edit images based on source and target prompts
-
Whisper Speaker Diarization
π£331Generate speakerβlabeled transcripts from audio files
-
nvidia/Llama-3.1-Nemotron-70B-Instruct-HF
Text Generation β’ 71B β’ Updated β’ 8.07k β’ 2.07k -
Selenium Screenshot Gradio
π¨24Generate a website screenshot from a URL
-
IllusionDiffusion
π5.42kGenerate stunning high quality illusion artwork
-
Edge TTS Text To Speech
π1.25kGenerate speech audio from text with custom voice settings
-
MinerU Document Extraction Tools
π663Embedded MinerU document extraction demo
-
Florence 2
π845Generate captions, detections, and segmentations from images
-
Rag Community Tool Template
π10Find relevant text chunks from documents based on a query
-
MagicQuill
πͺΆ2.3kEdit images with sketches, colors, and text prompts
-
Qwen2.5 Coder Artifacts
π’1.73kGenerate and preview app code from a text description
-
Open Source Ai Year In Review 2024
π»552What happened in open-source AI this year, and whatβs next?
-
MV Adapter T2MV Anime
π106Generate anime-style multi-view images from texts
-
Sound AI SFX
π328SText to Audio(Sound SFX) Generator
-
AnyCoder
π3.3kGenerate code snippets with AI for web and app frameworks
-
Stable Diffusion 3.5 Large
π2.03kGenerate images with SD3.5
-
FLUX VisionReply
π¦54FLUX, Image to Texto to Image, VLM
-
PuLID-FLUX
π€2.11kGenerate customized images from text and reference photos
-
Echo Chatbot Gradio Discord Bot
β‘2 -
FLUX.1 [dev]
π₯9.5kGenerate images from text prompts
-
Leffa
π619Generate realistic person images with new clothes or poses
-
AniDoc
π§76Animation Sketches sequence Colorization
-
PhotoMaker V2
π·1.22kGenerate personalized portrait images from your photos and prompts
-
Chat With Janus-Pro-7B
π2.02kA unified multimodal understanding and generation model.
-
DeepSite v4
π³16.6kGenerate any application by Vibe Coding it
-
InfiniteYou-FLUX
πΈ1.1kFlexible Photo Recrafting While Preserving Your Identity
-
Hi3DGen
π’702High-fidelity 3D Geometry Generation from single view image
-
DeepSite Gallery
π940Browse apps made with DeepSite
-
Kolors Virtual Try-On
π10.2kGenerate virtual tryβon images of a person wearing a chosen garment
-
Tabli
πΌ1Table of md, txt csv to image table parser
-
Realistic Text To Speech Unlimited
π₯1.76kFree Text-To-Speech generator with Emotion control (OpenAI)
-
Tar
π48Unified MLLM with Text-Aligned Representations
-
Reverse Face Search
π155Search Face Online
-
AutoPage
π12Generate project pages from research papers
-
Facetorch App
π43Facial expressions, 3D landmarks, embeddings, recognition.
-
Hello World
π€108One App to Rule Them All β 146 APIs, 81 emotions
-
See-through: Layer Decomposition
π140Generate layered PSD of anime characters from a single image
-
roop-unleashed
π»6FaceSwap tool_Image & Video
-
Face Swap App
π€42Swap faces in images and videos
-
NSFW Uncensored Photo
π49NSFW FLUX Uncensored photo 'Text & Imagery for AI Limits'
-
Kokoro TTS
β€3.42kUpgraded to v1.0!
-
Gemma4
π₯32Chat with Gemma 4 models fully in the browser via MediaPipe
-
Model Explorer
π48Explore and visualize machine learning model architectures
-
Smarter NPC
π€36Choose NPC action based on user input
-
Huggingfab
π¦30Generate 3D models from text descriptions
-
MiniMaxAI/MiniMax-M3
Image-Text-to-Text β’ 427B β’ Updated β’ 161k β’ β’ 1.44k