Inference Providers
Active filters: 8-bit
Text Generation
• 2B • Updated • 628
• 9
nvidia/DeepSeek-R1-0528-NVFP4
Text Generation
• 397B • Updated • 9.46k
• 45
nvidia/Qwen2.5-VL-7B-Instruct-NVFP4
Text Generation
• 5B • Updated • 5.53k
• 16
amd/gpt-oss-120b-w-mxfp4-a-fp8
59B • Updated • 13.2k
• 10
openai/gpt-oss-safeguard-20b
Text Generation
• 22B • Updated • 82.5k
• • 255
lmstudio-community/Qwen3-VL-8B-Instruct-MLX-8bit
Image-Text-to-Text
• Updated • 88.3k
• 6
Disty0/Chroma1-HD-SDNQ-uint4-svd-r32
Text-to-Image
• 5B • Updated • 371
• 1
EZCon/Huihui-Qwen3-VL-4B-Instruct-abliterated-8bit-mlx
Image-Text-to-Text
• 2B • Updated • 407
• 1
deng8470/FLUX.1-t5-xxl-HQQ-int8
5B • Updated • 6
• 1
235B • Updated • 1
RedHatAI/Qwen3-235B-A22B-NVFP4
Text Generation
• 136B • Updated • 1.46k
• 2
mingyi456/VibeVoice-7B-DF11
Text-to-Speech
• 13B • Updated • 30
• 3
johnsmith968530/Qwen-Qwen3-VL-8B-Thinking-MLX-8bit
Image-Text-to-Text
• Updated • 54
• 1
diffusers/FLUX.2-dev-bnb-4bit
Image-to-Image
• 17B • Updated • 24.6k
• 32
Disty0/Z-Image-Turbo-SDNQ-int8
Text-to-Image
• 6B • Updated • 4.03k
• 20
nvidia/Kimi-K2-Thinking-NVFP4
Text Generation
• 519B • Updated • 65.5k
• 33
Robotics
• 4B • Updated • 1
speakleash/Bielik-11B-v3.0-Instruct-MLX-8bit
Text Generation
• 11B • Updated • 285
• 3
MultiverseComputingCAI/HyperNova-60B
Text Generation
• 60B • Updated • 513
• 62
mlx-community/VulnLLM-R-7B-8bit
Text Generation
• 8B • Updated • 81
• 2
aydin99/FLUX.2-klein-4B-int8
Text-to-Image
• 4B • Updated • 740
• 13
lmstudio-community/GLM-4.7-Flash-MLX-8bit
Text Generation
• 30B • Updated • 203k
• 13
mlx-community/Qwen3-TTS-12Hz-1.7B-VoiceDesign-8bit
Text-to-Speech
• 0.8B • Updated • 2.18k
• 10
mlx-community/Qwen3-TTS-12Hz-1.7B-Base-8bit
Text-to-Speech
• 0.8B • Updated • 3.71k
• 12
mlx-community/Qwen3-TTS-12Hz-1.7B-CustomVoice-8bit
Text-to-Speech
• 0.8B • Updated • 3.66k
• 32
mlx-community/Qwen3-ASR-1.7B-8bit
0.8B • Updated • 3.57k
• 18
mlx-community/Qwen3-ASR-0.6B-8bit
0.4B • Updated • 220k
• 5
mlx-community/GLM-OCR-8bit
Image-to-Text
• 0.6B • Updated • 805
• 9
APMIC/Llama-Primus-Nemotron-70B-Instruct-nvfp4
36B • Updated • 75
• 1
MuXodious/LFM2.5-VL-1.6B-absolute-heresy-MPOA-mlx-8Bit
Image-Text-to-Text
• 0.7B • Updated • 40
• 1