Inference Providers
Active filters: 3-bit
Text Generation
• 7B • Updated • 1.18k
• 148
0xSero/GLM-5.3-500B-EXL3-3.0bpw
Text Generation
• 99B • Updated • 856
• 24
orcarouter/DeepSeek-V4.1-Flash-Uncensored-MLX
Image-Text-to-Text
• Updated • 39
Litwein/Qwen3.8-Flash-Next-REAP320-oQ3e-fp16-DWQ-MTP-Vision-MLX
Image-Text-to-Text
• 134B • Updated • 655
• 4
Brooooooklyn/Qwen3.5-9B-unsloth-mlx
Text Generation
• 9B • Updated • 594
• 15
drowzeys/keys-GLM-5.3-EXL3-Abliterated
Text Generation
• 165B • Updated • 184
• 12
Litwein/Qwen3.8-Flash-Next-REAP320-oQ3e-DWQ-MTP-Vision-MLX
Image-Text-to-Text
• 134B • Updated • 755
• 4
ukisai/Swift-1.5-3bit-MLX-TextOnly
Text Generation
• 27B • Updated • 279
• 2
MaziyarPanahi/Llama-3-Groq-8B-Tool-Use-GGUF
Text Generation
• 8B • Updated • 1.6k
• 10
numen-tech/Dolphin3.0-Llama3.1-8B-w3a16g40sym
Text Generation
• Updated • 1
RossAscends/24B-XortronCriminalComputingConfig-3bpw-EXL3
Text Generation
• 5B • Updated • 9
• 1
MaziyarPanahi/GLM-4.6V-Flash-GGUF
Text Generation
• 9B • Updated • 89.7k
• 7
RepublicOfKorokke/GLM-4.7-Flash-REAP-23B-A3B-oQ3.5
Text Generation
• 23B • Updated • 60
• 1
andrevp/Qwen3.6-35B-A3B-3bit-MLX
35B • Updated • 340
• 4
Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-3bit
Text Generation
• 27B • Updated • 207
• 2
majentik/gemma-4-12B-TurboQuant-MLX-3bit
Image-Text-to-Text
• 12B • Updated • 104
• 1
AtomicChat/DeepSeek-V4-Flash-0731-GGUF
Text Generation
• 284B • Updated • 2.63k
• 32
mastouri/GLM-5.2-colibri-E8-IQ3-with-int8-mtp
Text Generation
• Updated • 721
• 15
mickyba/Qwen3.8-27B-3bit-mlx
Image-Text-to-Text
• 27B • Updated • 285
• 1
leonsarmiento/Qwen3.8-27B-3bit-mtp-mlx
Image-Text-to-Text
• 28B • Updated • 2.92k
• 8
nathansutton/Qwen3.8-27B-UD-Q3_K_XL-DFlash2-MLX
Text Generation
• 27B • Updated • 5.34k
• 4
Litwein/Qwen3.8-Flash-Next-REAP320-oQ3e-DWQ-MTP-Vision-MTPLX
Image-Text-to-Text
• 81B • Updated • 1.02k
• 5
Litwein/Qwen3.8-Flash-Next-REAP320-oQ3e-fp16-DWQ-MTP-Vision-MTPLX
Image-Text-to-Text
• 81B • Updated • 706
• 2
NovaeonStudio/Qwen3.8-Flash-Next-Uncensored-oQ3e-fp16-mtp
Image-Text-to-Text
• 180B • Updated • 1.48k
• 2
kaitchup/Llama-2-7b-gptq-3bit
Text Generation
• Updated • 13
clibrain/Llama-2-7b-ft-instruct-es-gptq-3bit
Text Generation
• Updated • 12
• 3
clibrain/Llama-2-13b-ft-instruct-es-gptq-3bit
Text Generation
• Updated • 13
• 3
MiNeves-tops/opt-125m-gptq-3bit
Text Generation
• Updated • 15
Text Generation
• Updated • 12
LoneStriker/Yi-6B-200K-3.0bpw-h6-exl2
Text Generation
• Updated • 26