FP8 & NVFP4 quants of Qwen/Qwen3.8-27B. Reserved ahead of the upstream release — weights land as soon as the base model ships.
Huggin Fork
huginnfork
AI & ML interests
None yet
Recent Activity
new activity 11 days ago
huginnfork/Qwen3.8-27B-FP8:vllm: Vision Model missing warning new activity 14 days ago
huginnfork/Qwen3.8-27B-NVFP4A16:Will there be a GGUF version of the model? new activity 15 days ago
huginnfork/Qwen3.8-27B-NVFP4A16:Does not work with vllm 0.27.1Organizations
None yet
Tess-4-27B
FP8 & NVFP4 quants of migtissera/Tess-4-27B, a Qwen3.6-27B reasoning fine-tune. KLD is measured against bf16 Tess, so it isolates quant loss.
Qwen3.6-27B
FP8 & NVFP4 quants of Qwen/Qwen3.6-27B, plus the heretic-v2 uncensored line. Vision tower, SSM block and MTP head kept in bf16.
-
huginnfork/Qwen3.6-27B-NVFP4A16
Image-Text-to-Text • 20B • Updated • 166 • 1 -
huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp
Image-Text-to-Text • 28B • Updated • 27 • 4 -
huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp-FP8
Image-Text-to-Text • 28B • Updated • 1.11k • 5 -
huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp-NVFP4A16
Image-Text-to-Text • 20B • Updated • 106 • 2
Ornith-1.0-35B
FP8 & NVFP4 quants of the 256-expert Ornith-1.0-35B MoE. Routers, SSM block and vision tower stay bf16. Serve in vLLM or SGLang.
ThinkingCap-Qwen3.6-27B
FP8 quants of bottlecapai/ThinkingCap-Qwen3.6-27B. bf16 SSM/attention layout, and unlike the upstream FP8 it loads in transformers on Blackwell.
Qwen3.8-27B
FP8 & NVFP4 quants of Qwen/Qwen3.8-27B. Reserved ahead of the upstream release — weights land as soon as the base model ships.
Ornith-1.0-35B
FP8 & NVFP4 quants of the 256-expert Ornith-1.0-35B MoE. Routers, SSM block and vision tower stay bf16. Serve in vLLM or SGLang.
Tess-4-27B
FP8 & NVFP4 quants of migtissera/Tess-4-27B, a Qwen3.6-27B reasoning fine-tune. KLD is measured against bf16 Tess, so it isolates quant loss.
ThinkingCap-Qwen3.6-27B
FP8 quants of bottlecapai/ThinkingCap-Qwen3.6-27B. bf16 SSM/attention layout, and unlike the upstream FP8 it loads in transformers on Blackwell.
Qwen3.6-27B
FP8 & NVFP4 quants of Qwen/Qwen3.6-27B, plus the heretic-v2 uncensored line. Vision tower, SSM block and MTP head kept in bf16.
-
huginnfork/Qwen3.6-27B-NVFP4A16
Image-Text-to-Text • 20B • Updated • 166 • 1 -
huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp
Image-Text-to-Text • 28B • Updated • 27 • 4 -
huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp-FP8
Image-Text-to-Text • 28B • Updated • 1.11k • 5 -
huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp-NVFP4A16
Image-Text-to-Text • 20B • Updated • 106 • 2