Inference Providers
Active filters: mlx
TheBlueObserver/Llama-3.2-3B-Instruct-MLX-104ce
Text Generation
• 0.4B • Updated • 8
cubeus/Behemoth-123B-v1.2-Q8-mlx
34B • Updated • 16
mlx-community/EXAONE-3.5-7.8B-Instruct-3bit
Text Generation
• 1.0B • Updated • 90
TheBlueObserver/Falcon3-7B-Instruct-MLX-393a7
Text Generation
• 2B • Updated • 5
mlx-community/QwQ-32B-Preview-2bit
Text Generation
• 3B • Updated • 10
• 1
TheBlueObserver/Llama-3.2-3B-Instruct-MLX-196c8
Text Generation
• 0.5B • Updated • 7
mlx-community/aya-expanse-8b-4bit-mlx
Text Generation
• 1B • Updated • 21
TheBlueObserver/Llama-3.2-3B-Instruct-MLX-8777b
Text Generation
• 0.7B • Updated • 10
TheBlueObserver/Llama-3.2-3B-Instruct-MLX-393a7
Text Generation
• 0.9B • Updated • 8
TheBlueObserver/gemma-2-27b-it-MLX-393a7
Text Generation
• 8B • Updated • 17
mlx-community/Llama-3.1-Swallow-8B-Instruct-v0.3-8bit
Text Generation
• 2B • Updated • 11
mlx-community/Llama-3.1-Swallow-8B-Instruct-v0.3-4bit
Text Generation
• 1B • Updated • 27
ljnlonoljpiljm/florence-2-base-test-od-ft-mlx
Image-Text-to-Text
• 0.3B • Updated • 7
TheBlueObserver/Qwen2.5-Coder-32B-Instruct-MLX-8777b
Text Generation
• 7B • Updated • 17
TheBlueObserver/Qwen2.5-Coder-32B-Instruct-MLX-393a7
Text Generation
• 9B • Updated • 13
reach-vb/Qwen2.5-0.5B-Instruct-Q3-mlx
Text Generation
• 61.8M • Updated • 15
mlx-community/QVQ-72B-Preview-4bit
Image-Text-to-Text
• 11B • Updated • 24
• 7
mlx-community/QVQ-72B-Preview-3bit
Image-Text-to-Text
• 9B • Updated • 24
• 5
mlx-community/QVQ-72B-Preview-6bit
Image-Text-to-Text
• 16B • Updated • 35
• 2
mlx-community/QVQ-72B-Preview-8bit
Image-Text-to-Text
• 21B • Updated • 24
• 3
mlx-community/QVQ-72B-Preview-bf16
Image-Text-to-Text
• 73B • Updated • 27
• 3
BillSYZhang/gte-Qwen2-7B-instruct-Q4-mlx
Sentence Similarity
• 1B • Updated • 11
1B • Updated • 49
cnfusion/Microsoft_Phi-4-mlx-6bit
Text Generation
• 3B • Updated • 23
qkC8KADS/QWQ-Rombos-ties-TEST2-mlx_4
Text Generation
• 5B • Updated • 8
ubaitur5/Qwen2.5-0.5B-Instruct-Q3-mlx
Text Generation
• 61.8M • Updated • 29
mlx-community/Qwen2.5-14B-Instruct-3bit
Text Generation
• 2B • Updated • 78
• 1
mlx-community/nanoLLaVA-1.5-3bit
0.1B • Updated • 8
mlx-community/nanoLLaVA-1.5-6bit
0.2B • Updated • 7
ljnlonoljpiljm/florence-2-base-ft-objects-mlx
Image-Text-to-Text
• 0.3B • Updated • 7