John Ho PRO
AI & ML interests
Recent Activity
Organizations
- Running19
Quant
💻19Create interactive web apps with Python in minutes
- RunningFeatured94
LFM2 WebGPU – In-browser tool calling
🛠94In-browser tool calling, powered by Transformers.js
- Running3.31k
AnyCoder
🏆3.31kGenerate code snippets with AI for web and app frameworks
- RunningFeatured65
Privacy Filter WebGPU
🕵65PII detection and text masking in your browser
- Running on ZeroAgentsFeatured63
LightGlue
↔63LightGlue demo
- Running on ZeroMCPFeatured36
Qwen3 VL HF Demo
🔥36Object Detection, Visual Grounding, Keypoint Detection
-
prithivMLmods/MetaCLIP-2-Age-Range-Estimator
Image Classification • 21.7M • Updated • 113 • 7 - RunningFeatured765
Remove Background Web
🖼765In-browser background removal
- RunningAgents23
AI Video Editor
🏞23Create videos with FFMPEG + Qwen2.5-Coder
-
Searchium-ai/clip4clip-webvid150k
Text-to-Video • 0.2B • Updated • 201 • 47 - RunningFeatured449
FastVLM WebGPU
🍎449Real-time video captioning powered by FastVLM
- PausedAgentsFeatured36
AudioRag Demo
🎵36Search audio for relevant chunks
- Running on ZeroAgentsFeatured482
Parakeet-TDT-0.6b-V2
482Transcribe audio files with timestamps and downloadable subtitles
- Running on ZeroAgents55
Fast Whisper Turbo
⚡55Ultra-fast Whisper Turbo inference ⚡
-
openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.64M • • 3.38k - Running on ZeroAgentsFeatured344
Realtime Whisper Turbo
🤯344Realtime implementation of Whisper large turbo
- Running on T4Agents154
RF-DETR
🔥154SOTA real-time object detection model
- Running on CPU UpgradeAgents50
YOLO ARENA
🏟50compare performance of top object detectors
- RunningAgents24
SAM2 Video Predictor
🔥24Segment and track objects in videos
- Running on ZeroAgentsFeatured123
VLM Object Understanding
🦀123Explore object detection, visual grounding, keypoint Detecti
- Running on ZeroAgentsFeatured111
Qwen2 VL Localization
📉111Detect objects in images using text prompts
- Build errorAgentsFeatured160
Seed1.5 VL
🚀160Seed1.5-VL API Demo
- PausedAgents2
Vision Language SmolVLM2
🌍2Video + text to text with SmolVLM2
- Running on ZeroAgentsFeatured144
Gemma 3n E4B It
⚡144Chat with an AI that understands text, images, audio, and video
- Runtime errorAgents9
Cantonese TTS Text To Speech
👁9Generate Cantonese speech from text
- Runtime errorAgents4
Cantonese TTS Playground
🔥4Generate speech from Cantonese text using selected or custom voice
- Running on ZeroAgentsFeatured1.79k
Dia 1.6B
👯1.79kGenerate realistic dialogue from a script, using Dia!
- PausedAgentsFeatured81
Daily Paper Podcast
🎙81Generates a podcast about today's top trending paper.
- RunningAgentsFeatured255
PaddleOCR-VL Online Demo
📈255Extract text, tables, formulas, and charts from images
- Running on ZeroAgentsFeatured453
DeepSeek OCR Demo
🆘453An interactive demo for the DeepSeek-OCR model.
- Running on ZeroAgentsFeatured126
LightOnOCR 2 1B Demo
🐨126Extract text and layout from images or PDF documents
- Running on ZeroMCPFeatured143
Multimodal OCR2
💻143FireRed / Nanonets / Monkey / Thyme / Typhoon / SmolDocling
- Build error51
Quant
💻51Display interactive data visualizations and apps
- RunningFeatured49
Porting nanochat to Transformers: an AI modeling history lesson
📝49Read an AI article with theme toggle and PDF download
- Running on CPU UpgradeFeatured3.3k
The Smol Training Playbook
📚3.3kThe secrets to building world-class LLMs
- Running on ZeroAgentsFeatured849
Florence 2
📉849Generate captions, detections, and segmentations from images
- Running on ZeroAgentsFeatured519
Florence2 + SAM2
🔥519Segment objects in images or videos using text prompts
- SleepingAgentsFeatured120
SAM2 Video Predictor
🔥120Generate object masks and masked video from your MP4
- RunningAgents24
SAM2 Video Predictor
🔥24Segment and track objects in videos
-
EvanZhouDev/open-genmoji
Text-to-Image • Updated • 97 • • 68 - Running on ZeroAgentsFeatured682
ACE Step
😻682A Step Towards Music Generation Foundation Model
- Running on ZeroAgentsFeatured606
DreamO
🐨606A Unified Framework for Image Customization
- Running on ZeroAgentsFeatured1.01k
Tile Upscaler
🚀1.01kEnhance and upscale images with tile‑based AI control
- Running on ZeroAgentsFeatured1.45k
EasyControl Ghibli
🦀1.45kNew Ghibli EasyControl model is now released!!
-
akiyamasho/AnimeBackgroundGAN-Miyazaki
Image-to-Image • Updated • 25 - Running on ZeroAgents249
Ghibli Multilingual Text-Rendering
🦀249Elevating Ghibli-style AI art beyond ChatGPT's capabilities.
- Build errorMCP46
EasyControl Ghibli
🦀46New Ghibli EasyControl model is now released!!
- RunningAgentsFeatured255
PaddleOCR-VL Online Demo
📈255Extract text, tables, formulas, and charts from images
- Running on ZeroAgentsFeatured453
DeepSeek OCR Demo
🆘453An interactive demo for the DeepSeek-OCR model.
- Running on ZeroAgentsFeatured126
LightOnOCR 2 1B Demo
🐨126Extract text and layout from images or PDF documents
- Running on ZeroMCPFeatured143
Multimodal OCR2
💻143FireRed / Nanonets / Monkey / Thyme / Typhoon / SmolDocling
- Running19
Quant
💻19Create interactive web apps with Python in minutes
- RunningFeatured94
LFM2 WebGPU – In-browser tool calling
🛠94In-browser tool calling, powered by Transformers.js
- Running3.31k
AnyCoder
🏆3.31kGenerate code snippets with AI for web and app frameworks
- RunningFeatured65
Privacy Filter WebGPU
🕵65PII detection and text masking in your browser
- Running on ZeroAgentsFeatured63
LightGlue
↔63LightGlue demo
- Running on ZeroMCPFeatured36
Qwen3 VL HF Demo
🔥36Object Detection, Visual Grounding, Keypoint Detection
-
prithivMLmods/MetaCLIP-2-Age-Range-Estimator
Image Classification • 21.7M • Updated • 113 • 7 - RunningFeatured765
Remove Background Web
🖼765In-browser background removal
- Build error51
Quant
💻51Display interactive data visualizations and apps
- RunningFeatured49
Porting nanochat to Transformers: an AI modeling history lesson
📝49Read an AI article with theme toggle and PDF download
- Running on CPU UpgradeFeatured3.3k
The Smol Training Playbook
📚3.3kThe secrets to building world-class LLMs
- RunningAgents23
AI Video Editor
🏞23Create videos with FFMPEG + Qwen2.5-Coder
-
Searchium-ai/clip4clip-webvid150k
Text-to-Video • 0.2B • Updated • 201 • 47 - RunningFeatured449
FastVLM WebGPU
🍎449Real-time video captioning powered by FastVLM
- PausedAgentsFeatured36
AudioRag Demo
🎵36Search audio for relevant chunks
- Running on ZeroAgentsFeatured482
Parakeet-TDT-0.6b-V2
482Transcribe audio files with timestamps and downloadable subtitles
- Running on ZeroAgents55
Fast Whisper Turbo
⚡55Ultra-fast Whisper Turbo inference ⚡
-
openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.64M • • 3.38k - Running on ZeroAgentsFeatured344
Realtime Whisper Turbo
🤯344Realtime implementation of Whisper large turbo
- Running on ZeroAgentsFeatured849
Florence 2
📉849Generate captions, detections, and segmentations from images
- Running on ZeroAgentsFeatured519
Florence2 + SAM2
🔥519Segment objects in images or videos using text prompts
- SleepingAgentsFeatured120
SAM2 Video Predictor
🔥120Generate object masks and masked video from your MP4
- RunningAgents24
SAM2 Video Predictor
🔥24Segment and track objects in videos
- Running on T4Agents154
RF-DETR
🔥154SOTA real-time object detection model
- Running on CPU UpgradeAgents50
YOLO ARENA
🏟50compare performance of top object detectors
- RunningAgents24
SAM2 Video Predictor
🔥24Segment and track objects in videos
- Running on ZeroAgentsFeatured123
VLM Object Understanding
🦀123Explore object detection, visual grounding, keypoint Detecti
- Running on ZeroAgentsFeatured111
Qwen2 VL Localization
📉111Detect objects in images using text prompts
- Build errorAgentsFeatured160
Seed1.5 VL
🚀160Seed1.5-VL API Demo
- PausedAgents2
Vision Language SmolVLM2
🌍2Video + text to text with SmolVLM2
- Running on ZeroAgentsFeatured144
Gemma 3n E4B It
⚡144Chat with an AI that understands text, images, audio, and video
-
EvanZhouDev/open-genmoji
Text-to-Image • Updated • 97 • • 68 - Running on ZeroAgentsFeatured682
ACE Step
😻682A Step Towards Music Generation Foundation Model
- Running on ZeroAgentsFeatured606
DreamO
🐨606A Unified Framework for Image Customization
- Running on ZeroAgentsFeatured1.01k
Tile Upscaler
🚀1.01kEnhance and upscale images with tile‑based AI control
- Runtime errorAgents9
Cantonese TTS Text To Speech
👁9Generate Cantonese speech from text
- Runtime errorAgents4
Cantonese TTS Playground
🔥4Generate speech from Cantonese text using selected or custom voice
- Running on ZeroAgentsFeatured1.79k
Dia 1.6B
👯1.79kGenerate realistic dialogue from a script, using Dia!
- PausedAgentsFeatured81
Daily Paper Podcast
🎙81Generates a podcast about today's top trending paper.
- Running on ZeroAgentsFeatured1.45k
EasyControl Ghibli
🦀1.45kNew Ghibli EasyControl model is now released!!
-
akiyamasho/AnimeBackgroundGAN-Miyazaki
Image-to-Image • Updated • 25 - Running on ZeroAgents249
Ghibli Multilingual Text-Rendering
🦀249Elevating Ghibli-style AI art beyond ChatGPT's capabilities.
- Build errorMCP46
EasyControl Ghibli
🦀46New Ghibli EasyControl model is now released!!