Running on CPU Upgrade Agents Featured 1.41k Open ASR Leaderboard 🏆 1.41k Explore and compare speech‑recognition model benchmarks
google/gemma-4-E4B-it-qat-q4_0-unquantized-assistant Any-to-Any • 78.8M • Updated 1 day ago • 5.16k • 12
Gemma 4 Assistant GGUF Collection Gemma 4 MTP assistant drafters as GGUF (F16/Q8_0/Q5_K_M/Q4_K_M/Q4_K_S). Speculative-decoding heads for the atomic-llama-cpp-turboquant fork. • 4 items • Updated May 7 • 13
Alexzander85/Qwen3.5-9B-Claude-4.6-Opus-Reasoning-Distilled-NVFP4-MLP-FP8KV Text Generation • 8B • Updated Mar 4 • 303 • 9
mconcat/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-NVFP4 Text Generation • 22B • Updated Mar 19 • 350 • 48