deepseek-ai/DeepSeek-V4-Flash-0731 Text Generation β’ 304B β’ Updated 1 day ago β’ 156k β’ β’ 1.55k
thinkingmachines/Inkling-Small Image-Text-to-Text β’ 266B β’ Updated 2 days ago β’ 6.84k β’ β’ 218
thinkingmachines/Inkling-Small-NVFP4 Image-Text-to-Text β’ 156B β’ Updated 3 days ago β’ 55.2k β’ 57
nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 Text Generation β’ 45B β’ Updated 26 days ago β’ 69.5k β’ 127
view post Post 4249 Weβre releasing new Qwen3.6 quants that run 2.5Γ faster on your GPU. β‘Qwen3.6-27B NVFP4 runs on 24GB VRAM.35B-A3B can hit 17,561 tok/s (B200).We also improved accuracy, tool calling, agent use, and looping.Qwen3.6 NVFP4: https://huggingface.co/collections/unsloth/nvfp4Guide: https://unsloth.ai/docs/models/qwen3.6#nvfp4 See translation 1 reply Β· π 15 15 π₯ 11 11 π€ 1 1 + Reply
unsloth/Qwen3.6-35B-A3B-NVFP4-Fast Image-Text-to-Text β’ 22B β’ Updated 21 days ago β’ 406k β’ 94