Daniel Rosehill
AI & ML interests
Recent Activity
Organizations
-
oruk/orukeet
Automatic Speech Recognition • 0.6B • Updated • 21.4k • 78 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 7.54M • 3.88k -
mlx-community/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 1.85M • 58 -
handy-computer/parakeet-unified-en-0.6b-gguf
Automatic Speech Recognition • 0.6B • Updated • 1.54M • 9
-
netease-youdao/Confucius4-R2T2
Automatic Speech Recognition • 2B • Updated • 4.93k • 385 -
oruk/orukeet
Automatic Speech Recognition • 0.6B • Updated • 21.4k • 78 -
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 572k • • 1.15k -
Qwen/Qwen3-ASR-1.7B
Automatic Speech Recognition • 2B • Updated • 2.04M • • 1.12k
-
st1992/bert-restore-punctuation
Token Classification • Updated • 9 -
superwhisper/s1-mini
Text Generation • 0.8B • Updated • 5.92k • 357 -
oliverguhr/fullstop-punctuation-multilang-large
Token Classification • 0.6B • Updated • 238k • • 180 -
superwhisper/s1-mini-GGUF
Text Generation • 0.8B • Updated • 194k • 38
-
zai-org/GLM-ASR-Nano-2512
Automatic Speech Recognition • 2B • Updated • 62.3k • 388 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 35.1k • 463 -
microsoft/Phi-4-multimodal-instruct
Automatic Speech Recognition • 6B • Updated • 270k • 1.61k -
facebook/omniASR-LLM-7B
Automatic Speech Recognition • Updated • 39
-
futo-org/acft-whisper-tiny
Automatic Speech Recognition • 57.7M • Updated • 13 • 2 -
futo-org/acft-whisper-small.en
Automatic Speech Recognition • 0.3B • Updated • 10 • 2 -
futo-org/acft-whisper-base.en
Automatic Speech Recognition • 99.1M • Updated • 59 • 2 -
futo-org/acft-whisper-tiny.en
Automatic Speech Recognition • 57.7M • Updated • 19 • 1
-
openai/whisper-base
Automatic Speech Recognition • 72.6M • Updated • 1.62M • 290 -
openai/whisper-base.en
Automatic Speech Recognition • 72.6M • Updated • 275k • 45 -
onnx-community/whisper-base_timestamped
Automatic Speech Recognition • Updated • 18.2k • 32 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 1.68M • 37
- Runtime errorAgents
Baby Noise Cancellation Demo
👶AI-powered baby noise removal demo with STT comparison
- RunningAgents184
DeepFilterNet2
💩184Denoise your recordings and view spectrograms
- RunningAgents19
DeepFilterNet2 No File Size Limit
😻19Use DeepFilterNet2 to denoise audio no file size limit
-
benlehrburger/modern-architecture
Viewer • Updated • 1.09k • 659 • 4 - SleepingAgents2
ArchitectureClassifier
📈2Classify architectural styles in images
- RunningAgents17
Rocco Architecture Render
🚀17Generate interior and exterior designs from sketches
- SleepingAgents1
London Architecture
💻1Classify architectural styles in images
- Running368
SD Artists Browser
🤘368Explore artist styles and build SDXL prompts
- Running on ZeroMCP65
StyleAligned Transfer
🐠65Generate images in the style of a reference image
- PausedAgents17
StyleFeatureEditor
💻17Edit images with predefined styles or text prompts
- Runtime errorAgents12
Kontext Style LoRAs
🌍12Transform images using selected styles
- Running3
Pharmacology Graph Explorer
💊3Explore predicted drug‑target and disease links
- RunningAgents72
Medical Diagnosis
📉72Classify symptoms to diagnose health issues
- Running28
MediAI Medical AI Agent
🚀28AI-Powered Diagnosis & Treatment Assistant
- SleepingAgents
Lisdexamfetamine Split Dose Modeller
🚀Model split-dose protocols for lisdexamfetamine/Vyvanse
- Running on ZeroAgentsFeatured2.32k
MagicQuill
🪶2.32kEdit images with sketches, colors, and text prompts
- Build errorAgents20
AutoPR
🚀20Generate a Twitter or Xiaohongshu post from a research PDF
- Running157
Reverse Face Search
📉157Search Face Online
- Runtime errorAgents16
AI STORYTELLER
🏢16Generate a video from a story
-
dots-studio/dots.ocr
Image-Text-to-Text • 3B • Updated • 972k • 1.33k - Runtime error15
Ui Rev Doc Model
😻15Analysis of data on an invoice
- PausedAgentsFeatured145
Deepdoctection
🏃145Convert PDFs and images to structured text and layout data
- Running13
Docsifer
📚13Convert documents into clean, LLM-ready Markdown.
-
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 6.25M • • 7.87k -
Qwen/Qwen3-Omni-30B-A3B-Instruct
Any-to-Any • 35B • Updated • 639k • 1.01k -
moonshotai/Kimi-K2-Instruct-0905
Text Generation • 1T • Updated • 58.2k • • 788 -
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 2.32M • 3.54k
-
Qwen/Qwen3-VL-8B-Instruct
Image-Text-to-Text • 9B • Updated • 19.7M • • 1.14k -
zai-org/GLM-4.6
Text Generation • 357B • Updated • 17.8k • • 1.24k -
Qwen/Qwen3-VL-8B-Thinking
Image-Text-to-Text • 9B • Updated • 156k • • 225 -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 6.25M • • 7.87k
-
danielrosehill/Shakespearean-Text-Transformation-Prompts
Viewer • Updated • 1 • 105 -
danielrosehill/Speech-To-Text-System-Prompts-2
Viewer • Updated • 2 • 306 • 1 - SleepingAgents
System Prompt Reformatter
📚Reformats system prompts in the 2nd person and other edits
- SleepingAgents
BLUF Email Formatter
📧Format emails with clear subject lines and summaries
- Sleeping
Max Output Tokens Analysis
📊Display max output tokens for models over time
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
- Running
Single Shot Brevity Training
📈Using one example to train an LLM for informational brevity
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
modularai/Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 8.29k • 17 -
MaziyarPanahi/WizardLM-2-7B-GGUF
Text Generation • 7B • Updated • 151k • 83 -
MaziyarPanahi/mathstral-7B-v0.1-GGUF
Text Generation • 7B • Updated • 150k • 8 -
MaziyarPanahi/phi-4-GGUF
Text Generation • 15B • Updated • 152k • 10
-
nvidia/parakeet-tdt-0.6b-v2
Automatic Speech Recognition • Updated • 191k • 1.55k -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 37.3k • 171 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 35.1k • 463 -
facebook/omniASR-W2V-1B
Automatic Speech Recognition • Updated • 6
-
nvidia/parakeet-tdt-1.1b
Automatic Speech Recognition • Updated • 2.01k • 139 -
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 572k • • 1.15k -
nvidia/nemotron-3.5-asr-streaming-0.6b
Automatic Speech Recognition • 0.6B • Updated • 829k • • 1.13k -
handy-computer/parakeet-tdt-1.1b-gguf
Automatic Speech Recognition • 1B • Updated • 23.1k • 1
-
imvladikon/wav2vec2-large-xlsr-53-hebrew
Automatic Speech Recognition • 0.3B • Updated • 24.4k • 8 -
Mizurodp/wav2vec2-large-xls-r-300m-hebrew-colab
Automatic Speech Recognition • Updated • 22 • 1 -
imvladikon/wav2vec2-xls-r-300m-lm-hebrew
Automatic Speech Recognition • 0.3B • Updated • 73 • 4 -
imvladikon/wav2vec2-xls-r-1b-hebrew
Automatic Speech Recognition • 1.0B • Updated • 464 • 2
-
danielrosehill/daniel_whisper_finetune_large_v3_turbo_v2
Automatic Speech Recognition • 0.8B • Updated • 23 -
danielrosehill/daniel_whisper_finetune_medium_v2
Automatic Speech Recognition • 0.8B • Updated • 19 -
danielrosehill/daniel_whisper_finetune_tiny_v2
Automatic Speech Recognition • 37.8M • Updated • 16 -
danielrosehill/daniel_whisper_finetune_base_v2
Automatic Speech Recognition • 72.6M • Updated • 18
-
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 572k • • 1.15k -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 37.3k • 171 -
facebook/seamless-m4t-v2-large
Automatic Speech Recognition • 2B • Updated • 305k • 1.01k -
facebook/wav2vec2-base-960h
Automatic Speech Recognition • 94.4M • Updated • 1.47M • 406
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 17 -
danielrosehill/Long-Prompt-Experiment
Viewer • Updated • 92 • 76 - RunningAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
-
pyannote/voice-activity-detection
Automatic Speech Recognition • Updated • 575k • 242 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 7.54M • 3.88k -
pyannote/overlapped-speech-detection
Automatic Speech Recognition • Updated • 5.39k • 64 -
pipecat-ai/smart-turn-v3
Voice Activity Detection • Updated • 204
-
unsloth/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 3.23k • 20 -
unsloth/whisper-small
Automatic Speech Recognition • 0.2B • Updated • 895 • 7 -
unsloth/CrisperWhisper
Automatic Speech Recognition • 2B • Updated • 115 • 16 -
unsloth/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.11k • 12
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 17 - RunningAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- Running
STT Comparison
🦀Comparing STT models against audio
- Running on ZeroMCPFeatured622
LatentSync
👄622Audio Conditioned LipSync with Latent Diffusion Models
- Build errorAgentsFeatured1.43k
SadTalker
😭1.43kGenerate a talking face video from an image and audio
- RunningAgents189
Gradio Lipsync Wav2lip
👄189Generate lip‑synced video from a face image and audio
- RunningAgents71
Wav2lip Gpu
🌍71Create a talking‑head video from a photo and audio
- RunningAgents447
Remove Silence From Audio
🦀447Remove Silence From Audio
- Running on ZeroAgents410
Audio🔹Separator
🏃410Vocal and background audio separator
- Running on ZeroAgentsFeatured330
Audio Editing
🎧330Edit audios with text prompts
- Running on ZeroAgents478
Resemble Enhance
🚀478Enhance your audio with denoising and quality boost
-
tencent/HunyuanImage-3.0
Text-to-Image • 83B • Updated • 3.32k • • 1.13k -
black-forest-labs/FLUX.1-schnell
Text-to-Image • 12B • Updated • 520k • • 6k -
stabilityai/stable-diffusion-xl-base-1.0
Text-to-Image • 3B • Updated • 3.51M • • 8.22k -
Qwen/Qwen-Image
Text-to-Image • 20B • Updated • 299k • • 2.66k
- Running on ZeroAgents195
PSHuman
🏃195PHOTOREALISTIC HUMAN RECONSTRUCTION w/ CROSS-SCALE DIFF
- Runtime errorAgents11
Pifuhd
🐠11Generate 3D human models from images
- Running on ZeroAgents10
HumanWild
⚡10Generate 3D human reconstructions from images
- Runtime error51
HSMR
💀51Convert images of humans to biomechanically accurate 3D skeletons
- Running on ZeroAgentsFeatured1.67k
Expression Editor
🐨1.67kQuickly edit the expression of a face
- Running on ZeroAgentsFeatured1.54k
InstructPix2Pix
🚀1.54kEdit images using text instructions
-
Qwen/Qwen-Image-Edit-2509
Image-to-Image • 20B • Updated • 475k • • 1.25k -
Qwen/Qwen-Image
Text-to-Image • 20B • Updated • 299k • • 2.66k
- Running on ZeroAgents3.8k
Live Portrait
🤪3.8kApply the motion of a video on a portrait
- Running on ZeroMCPFeatured2.04k
Stable Video Diffusion 1.1
📺2.04kGenerate a short video from a single image
- Running on ZeroMCPFeatured1.62k
Wan2.1 Fast
🎥1.62kAnimate an image into a short video using a text prompt
- Running on ZeroMCP2.95k
Background Removal
🌘2.95kRemove backgrounds from images instantly
- Running on ZeroAgents2.98k
CLIP Interrogator
🕵2.98kGenerate detailed prompts from any image
- PausedAgents441
NoWatermark
⚡441Powerful Watermark Removal API
- RunningAgents141
Vectorizer AI
🌍141Convert images to SVG vectors with customizable settings
- Running on CPU UpgradeAgents46
Hebrew LLM Leaderboard
🥇46Explore LLM benchmark leaderboard with searchable filters
- SleepingAgents
Hebrew GPT Neo - Science Fiction and Fantasy
🧙Generate Hebrew text for science fiction and fantasy stories
- RunningAgents
מחולל נונסנס רובושאול
🤖מחולל משפטים בסגנון ״חיות כיס״
- Build errorAgents
Hebrew Sentiment
😻
- Running on CPU UpgradeAgentsFeatured1.47k
Open ASR Leaderboard
🏆1.47kCompare ASR model WER and speed across languages and datasets
- RunningAgents34
Hebrew Transcription Leaderboard
🥇34Benchmarking Hebrew Speech-to-Text Models
- RunningAgents453
Agent Leaderboard
💬453Ranking of LLMs for agentic tasks
-
oruk/orukeet
Automatic Speech Recognition • 0.6B • Updated • 21.4k • 78 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 7.54M • 3.88k -
mlx-community/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 1.85M • 58 -
handy-computer/parakeet-unified-en-0.6b-gguf
Automatic Speech Recognition • 0.6B • Updated • 1.54M • 9
-
netease-youdao/Confucius4-R2T2
Automatic Speech Recognition • 2B • Updated • 4.93k • 385 -
oruk/orukeet
Automatic Speech Recognition • 0.6B • Updated • 21.4k • 78 -
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 572k • • 1.15k -
Qwen/Qwen3-ASR-1.7B
Automatic Speech Recognition • 2B • Updated • 2.04M • • 1.12k
-
st1992/bert-restore-punctuation
Token Classification • Updated • 9 -
superwhisper/s1-mini
Text Generation • 0.8B • Updated • 5.92k • 357 -
oliverguhr/fullstop-punctuation-multilang-large
Token Classification • 0.6B • Updated • 238k • • 180 -
superwhisper/s1-mini-GGUF
Text Generation • 0.8B • Updated • 194k • 38
-
nvidia/parakeet-tdt-1.1b
Automatic Speech Recognition • Updated • 2.01k • 139 -
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 572k • • 1.15k -
nvidia/nemotron-3.5-asr-streaming-0.6b
Automatic Speech Recognition • 0.6B • Updated • 829k • • 1.13k -
handy-computer/parakeet-tdt-1.1b-gguf
Automatic Speech Recognition • 1B • Updated • 23.1k • 1
-
imvladikon/wav2vec2-large-xlsr-53-hebrew
Automatic Speech Recognition • 0.3B • Updated • 24.4k • 8 -
Mizurodp/wav2vec2-large-xls-r-300m-hebrew-colab
Automatic Speech Recognition • Updated • 22 • 1 -
imvladikon/wav2vec2-xls-r-300m-lm-hebrew
Automatic Speech Recognition • 0.3B • Updated • 73 • 4 -
imvladikon/wav2vec2-xls-r-1b-hebrew
Automatic Speech Recognition • 1.0B • Updated • 464 • 2
-
zai-org/GLM-ASR-Nano-2512
Automatic Speech Recognition • 2B • Updated • 62.3k • 388 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 35.1k • 463 -
microsoft/Phi-4-multimodal-instruct
Automatic Speech Recognition • 6B • Updated • 270k • 1.61k -
facebook/omniASR-LLM-7B
Automatic Speech Recognition • Updated • 39
-
danielrosehill/daniel_whisper_finetune_large_v3_turbo_v2
Automatic Speech Recognition • 0.8B • Updated • 23 -
danielrosehill/daniel_whisper_finetune_medium_v2
Automatic Speech Recognition • 0.8B • Updated • 19 -
danielrosehill/daniel_whisper_finetune_tiny_v2
Automatic Speech Recognition • 37.8M • Updated • 16 -
danielrosehill/daniel_whisper_finetune_base_v2
Automatic Speech Recognition • 72.6M • Updated • 18
-
nvidia/parakeet-tdt-0.6b-v3
Automatic Speech Recognition • 0.6B • Updated • 572k • • 1.15k -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 37.3k • 171 -
facebook/seamless-m4t-v2-large
Automatic Speech Recognition • 2B • Updated • 305k • 1.01k -
facebook/wav2vec2-base-960h
Automatic Speech Recognition • 94.4M • Updated • 1.47M • 406
-
futo-org/acft-whisper-tiny
Automatic Speech Recognition • 57.7M • Updated • 13 • 2 -
futo-org/acft-whisper-small.en
Automatic Speech Recognition • 0.3B • Updated • 10 • 2 -
futo-org/acft-whisper-base.en
Automatic Speech Recognition • 99.1M • Updated • 59 • 2 -
futo-org/acft-whisper-tiny.en
Automatic Speech Recognition • 57.7M • Updated • 19 • 1
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 17 -
danielrosehill/Long-Prompt-Experiment
Viewer • Updated • 92 • 76 - RunningAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
-
pyannote/voice-activity-detection
Automatic Speech Recognition • Updated • 575k • 242 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 7.54M • 3.88k -
pyannote/overlapped-speech-detection
Automatic Speech Recognition • Updated • 5.39k • 64 -
pipecat-ai/smart-turn-v3
Voice Activity Detection • Updated • 204
-
unsloth/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 3.23k • 20 -
unsloth/whisper-small
Automatic Speech Recognition • 0.2B • Updated • 895 • 7 -
unsloth/CrisperWhisper
Automatic Speech Recognition • 2B • Updated • 115 • 16 -
unsloth/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.11k • 12
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
danielrosehill/Podcast-ASR-Evaluation
Viewer • Updated • 27 • 17 - RunningAgents
Podcast ASR Evaluation
🎙ASR benchmark comparing local and cloud models
- Running
STT Comparison
🦀Comparing STT models against audio
-
openai/whisper-base
Automatic Speech Recognition • 72.6M • Updated • 1.62M • 290 -
openai/whisper-base.en
Automatic Speech Recognition • 72.6M • Updated • 275k • 45 -
onnx-community/whisper-base_timestamped
Automatic Speech Recognition • Updated • 18.2k • 32 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 1.68M • 37
- Runtime errorAgents
Baby Noise Cancellation Demo
👶AI-powered baby noise removal demo with STT comparison
- RunningAgents184
DeepFilterNet2
💩184Denoise your recordings and view spectrograms
- RunningAgents19
DeepFilterNet2 No File Size Limit
😻19Use DeepFilterNet2 to denoise audio no file size limit
- Running on ZeroMCPFeatured622
LatentSync
👄622Audio Conditioned LipSync with Latent Diffusion Models
- Build errorAgentsFeatured1.43k
SadTalker
😭1.43kGenerate a talking face video from an image and audio
- RunningAgents189
Gradio Lipsync Wav2lip
👄189Generate lip‑synced video from a face image and audio
- RunningAgents71
Wav2lip Gpu
🌍71Create a talking‑head video from a photo and audio
-
benlehrburger/modern-architecture
Viewer • Updated • 1.09k • 659 • 4 - SleepingAgents2
ArchitectureClassifier
📈2Classify architectural styles in images
- RunningAgents17
Rocco Architecture Render
🚀17Generate interior and exterior designs from sketches
- SleepingAgents1
London Architecture
💻1Classify architectural styles in images
- Running368
SD Artists Browser
🤘368Explore artist styles and build SDXL prompts
- Running on ZeroMCP65
StyleAligned Transfer
🐠65Generate images in the style of a reference image
- PausedAgents17
StyleFeatureEditor
💻17Edit images with predefined styles or text prompts
- Runtime errorAgents12
Kontext Style LoRAs
🌍12Transform images using selected styles
- Running3
Pharmacology Graph Explorer
💊3Explore predicted drug‑target and disease links
- RunningAgents72
Medical Diagnosis
📉72Classify symptoms to diagnose health issues
- Running28
MediAI Medical AI Agent
🚀28AI-Powered Diagnosis & Treatment Assistant
- SleepingAgents
Lisdexamfetamine Split Dose Modeller
🚀Model split-dose protocols for lisdexamfetamine/Vyvanse
- RunningAgents447
Remove Silence From Audio
🦀447Remove Silence From Audio
- Running on ZeroAgents410
Audio🔹Separator
🏃410Vocal and background audio separator
- Running on ZeroAgentsFeatured330
Audio Editing
🎧330Edit audios with text prompts
- Running on ZeroAgents478
Resemble Enhance
🚀478Enhance your audio with denoising and quality boost
- Running on ZeroAgentsFeatured2.32k
MagicQuill
🪶2.32kEdit images with sketches, colors, and text prompts
- Build errorAgents20
AutoPR
🚀20Generate a Twitter or Xiaohongshu post from a research PDF
- Running157
Reverse Face Search
📉157Search Face Online
- Runtime errorAgents16
AI STORYTELLER
🏢16Generate a video from a story
-
dots-studio/dots.ocr
Image-Text-to-Text • 3B • Updated • 972k • 1.33k - Runtime error15
Ui Rev Doc Model
😻15Analysis of data on an invoice
- PausedAgentsFeatured145
Deepdoctection
🏃145Convert PDFs and images to structured text and layout data
- Running13
Docsifer
📚13Convert documents into clean, LLM-ready Markdown.
-
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 6.25M • • 7.87k -
Qwen/Qwen3-Omni-30B-A3B-Instruct
Any-to-Any • 35B • Updated • 639k • 1.01k -
moonshotai/Kimi-K2-Instruct-0905
Text Generation • 1T • Updated • 58.2k • • 788 -
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 2.32M • 3.54k
-
tencent/HunyuanImage-3.0
Text-to-Image • 83B • Updated • 3.32k • • 1.13k -
black-forest-labs/FLUX.1-schnell
Text-to-Image • 12B • Updated • 520k • • 6k -
stabilityai/stable-diffusion-xl-base-1.0
Text-to-Image • 3B • Updated • 3.51M • • 8.22k -
Qwen/Qwen-Image
Text-to-Image • 20B • Updated • 299k • • 2.66k
-
Qwen/Qwen3-VL-8B-Instruct
Image-Text-to-Text • 9B • Updated • 19.7M • • 1.14k -
zai-org/GLM-4.6
Text Generation • 357B • Updated • 17.8k • • 1.24k -
Qwen/Qwen3-VL-8B-Thinking
Image-Text-to-Text • 9B • Updated • 156k • • 225 -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 6.25M • • 7.87k
-
danielrosehill/Shakespearean-Text-Transformation-Prompts
Viewer • Updated • 1 • 105 -
danielrosehill/Speech-To-Text-System-Prompts-2
Viewer • Updated • 2 • 306 • 1 - SleepingAgents
System Prompt Reformatter
📚Reformats system prompts in the 2nd person and other edits
- SleepingAgents
BLUF Email Formatter
📧Format emails with clear subject lines and summaries
- Running on ZeroAgents195
PSHuman
🏃195PHOTOREALISTIC HUMAN RECONSTRUCTION w/ CROSS-SCALE DIFF
- Runtime errorAgents11
Pifuhd
🐠11Generate 3D human models from images
- Running on ZeroAgents10
HumanWild
⚡10Generate 3D human reconstructions from images
- Runtime error51
HSMR
💀51Convert images of humans to biomechanically accurate 3D skeletons
- Running on ZeroAgentsFeatured1.67k
Expression Editor
🐨1.67kQuickly edit the expression of a face
- Running on ZeroAgentsFeatured1.54k
InstructPix2Pix
🚀1.54kEdit images using text instructions
-
Qwen/Qwen-Image-Edit-2509
Image-to-Image • 20B • Updated • 475k • • 1.25k -
Qwen/Qwen-Image
Text-to-Image • 20B • Updated • 299k • • 2.66k
- Running on ZeroAgents3.8k
Live Portrait
🤪3.8kApply the motion of a video on a portrait
- Running on ZeroMCPFeatured2.04k
Stable Video Diffusion 1.1
📺2.04kGenerate a short video from a single image
- Running on ZeroMCPFeatured1.62k
Wan2.1 Fast
🎥1.62kAnimate an image into a short video using a text prompt
- Running on ZeroMCP2.95k
Background Removal
🌘2.95kRemove backgrounds from images instantly
- Running on ZeroAgents2.98k
CLIP Interrogator
🕵2.98kGenerate detailed prompts from any image
- PausedAgents441
NoWatermark
⚡441Powerful Watermark Removal API
- RunningAgents141
Vectorizer AI
🌍141Convert images to SVG vectors with customizable settings
- Sleeping
Max Output Tokens Analysis
📊Display max output tokens for models over time
- SleepingAgents1
LLM Long Output Experiment (Code Generation)
📈1Evaluating max single output length of code gen LLMs
- Running
Single Shot Brevity Training
📈Using one example to train an LLM for informational brevity
- SleepingAgents
Local STT Eval One Sample
😻Single sample eval for WER on various Whisper models
-
modularai/Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 8.29k • 17 -
MaziyarPanahi/WizardLM-2-7B-GGUF
Text Generation • 7B • Updated • 151k • 83 -
MaziyarPanahi/mathstral-7B-v0.1-GGUF
Text Generation • 7B • Updated • 150k • 8 -
MaziyarPanahi/phi-4-GGUF
Text Generation • 15B • Updated • 152k • 10
- Running on CPU UpgradeAgents46
Hebrew LLM Leaderboard
🥇46Explore LLM benchmark leaderboard with searchable filters
- SleepingAgents
Hebrew GPT Neo - Science Fiction and Fantasy
🧙Generate Hebrew text for science fiction and fantasy stories
- RunningAgents
מחולל נונסנס רובושאול
🤖מחולל משפטים בסגנון ״חיות כיס״
- Build errorAgents
Hebrew Sentiment
😻
- Running on CPU UpgradeAgentsFeatured1.47k
Open ASR Leaderboard
🏆1.47kCompare ASR model WER and speed across languages and datasets
- RunningAgents34
Hebrew Transcription Leaderboard
🥇34Benchmarking Hebrew Speech-to-Text Models
- RunningAgents453
Agent Leaderboard
💬453Ranking of LLMs for agentic tasks
-
nvidia/parakeet-tdt-0.6b-v2
Automatic Speech Recognition • Updated • 191k • 1.55k -
ibm-granite/granite-speech-3.3-8b
Automatic Speech Recognition • 9B • Updated • 37.3k • 171 -
nvidia/canary-qwen-2.5b
Automatic Speech Recognition • 3B • Updated • 35.1k • 463 -
facebook/omniASR-W2V-1B
Automatic Speech Recognition • Updated • 6