Inference Providers
Active filters: 8-bit
Freaksterz/Qwen3.8-27B-SmoothQuant-W8A8-INT8
Image-Text-to-Text
• 27B • Updated • 21.2k
• 14
mlx-community/Qwen3.8-27B-MTP-8bit
Text Generation
• 0.1B • Updated • 11.8k
• 15
sakamakismile/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-NVFP4
Text Generation
• 27B • Updated • 47.1k
• 20
chimingw/Qwen3.8-27B-Uncensored-OrcaRouter-MLX-8bit
Image-Text-to-Text
• 9B • Updated • 14.1k
• 8
Jundot/Qwen3.8-27B-oQ8e-fp16-mtp
8B • Updated • 1.45k
• 3
HangGlidersRule/Darkstar-Qwen3.8-27B-Abliterated-ModelOpt-W4A16-NVFP4-Mixed-FP8
Text Generation
• 18B • Updated • 2.82k
• 3
cbert33/DeepSeek-V4-Flash-0731-abliterated-vision-v2
Image-Text-to-Text
• 304B • Updated • 437
• 3
Vontra/Qwen3.8-Flash-Next-MLX-oQ8-MTP
Image-Text-to-Text
• 56B • Updated • 1.55k
• 2
Vontra/Qwen3.8-Flash-Next-MLX-8bit-MTP
Image-Text-to-Text
• 57B • Updated • 2.51k
• 2
Blackfrost-AI/Qwen3.8-Flash-Next-DERISKED-NVFP4
Image-Text-to-Text
• 120B • Updated • 164
• 2
mazinb/Qwen3.8-Flash-Next-Uncensored-NVFP4
Image-Text-to-Text
• 120B • Updated • 458
• 3
RadixArk/GLM-5.3-Flash-NVFP4
Image-Text-to-Text
• 168B • Updated • 12
• 2
INCModel3/Qwen3.8-Flash-Next-MXFP4-Mixed-CT-AutoRound
Image-Text-to-Text
• 180B • Updated • 55
• 2
malaiwah/GLM-5.3-Flash-TR3-8bpw
Image-Text-to-Text
• 166B • Updated • 115
• 2
Blackfrost-Research/GLM-5.3-DERISKED-NVFP4
Text Generation
• 391B • Updated • 2
• 2
Text Generation
• 390B • Updated • 549
• 2
acyildirimer/Qwen3.8-27B-AutoQuant-NVFP4-FP8-5bit
Image-Text-to-Text
• 17B • Updated • 112
• 2
TheDrainFlorist/Qwen3.8-Flash-Next-VQ-2.1bpw
Text Generation
• 21B • Updated • 766
• 2
thebriangao/GLM-5.3-Flash-Uncensored-NVFP4
Image-Text-to-Text
• 321B • Updated • 78
• 2
primitive-ai/DeepSeek-V4-Flash-Vision-Exp-REAP-145B
Text Generation
• 146B • Updated • 202
• 2
MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF
Text Generation
• 7B • Updated • 172k
• 148
MaziyarPanahi/Llama-3.3-70B-Instruct-GGUF
Text Generation
• 71B • Updated • 171k
• 22
mlx-community/Mistral-Small-24B-Instruct-2501-8bit
7B • Updated • 79
• 4
ai-sage/GigaChat-20B-A3B-instruct-v1.5-int8
21B • Updated • 44
• 2
RedHatAI/DeepSeek-R1-Distill-Qwen-32B-quantized.w8a8
Text Generation
• 33B • Updated • 383
• 14
mlx-community/Qwen3-0.6B-8bit
Text Generation
• 0.2B • Updated • 99.6k
• 8
mxmcc/Llama-xLAM-2-8b-fc-r-mlx-8Bit
Text Generation
• 2B • Updated • 72
• 1
diffusers/FLUX.1-dev-bnb-4bit
Text-to-Image
• 6B • Updated • 2.27k
• 6
Text Generation
• 2B • Updated • 628
• 9
nvidia/DeepSeek-R1-0528-NVFP4
Text Generation
• 397B • Updated • 9.46k
• 45