Inference Providers
Active filters: mlx
mlx-community/Meta-Llama-3.1-70B-Instruct-bf16
Text Generation
• 71B • Updated • 146
• 3
mlx-community/Meta-Llama-3.1-70B-Instruct-4bit
Text Generation
• 11B • Updated • 5.03k
• 6
mlx-community/Meta-Llama-3.1-70B-Instruct-8bit
Text Generation
• 20B • Updated • 2.43k
• 4
mlx-community/Meta-Llama-3.1-70B-8bit
Text Generation
• 20B • Updated • 26
• 1
mlx-community/Meta-Llama-3.1-70B-4bit
Text Generation
• 11B • Updated • 39
• 1
mlx-community/Meta-Llama-3.1-70B-bf16
Text Generation
• 71B • Updated • 63
• 4
mlx-community/Meta-Llama-3.1-8B-bf16
Text Generation
• 8B • Updated • 29
• 1
mlx-community/Meta-Llama-3.1-8B-4bit
Text Generation
• 1B • Updated • 162
• 6
mlx-community/Meta-Llama-3.1-8B-8bit
Text Generation
• 2B • Updated • 24
context-labs/Meta-Llama-3.1-70B-4bit
Text Generation
• 11B • Updated • 19
• 1
context-labs/Meta-Llama-3.1-8B-4bit
Text Generation
• 1B • Updated • 7
mlx-community/Meta-Llama-3.1-405B-2bit
Text Generation
• 38B • Updated • 45
• 2
mlx-community/Mistral-Large-Instruct-2407-4bit
19B • Updated • 435
• 1
mlx-community/Mistral-Large-Instruct-2407-8bit
34B • Updated • 105
• 1
mlx-community/Mistral-Large-Instruct-2407-bf16
123B • Updated • 88
• 1
mlx-community/Dolphin-2.9.3-Mistral-Nemo-12b-bf16
12B • Updated • 100
• 2
mlx-community/Dolphin-2.9.3-Mistral-Nemo-12b-4bit
2B • Updated • 202
• 1
mlx-community/Dolphin-2.9.3-Mistral-Nemo-12b-8bit
3B • Updated • 66
• 1
sosoai/hansoldeco-Llama-3.1-8b-instruct-v0.1-mlx
1B • Updated • 7
mlx-community/Meta-Llama-3.1-405B-4bit
Text Generation
• 64B • Updated • 286
• 4
Text Generation
• 1B • Updated • 17
mlx-community/deepseek-vl-1.3b-chat-8bit
Image-Text-to-Text
• 0.6B • Updated • 21
• 1
mlx-community/deepseek-vl-7b-chat-8bit
Image-Text-to-Text
• 2B • Updated • 23
mlx-community/Llama-3.1-70B-Japanese-Instruct-2407-8bit
Text Generation
• 20B • Updated • 22
• 1
mlx-community/Meta-Llama-3.1-8B-Instruct-abliterated-4bit
Text Generation
• 1B • Updated • 188
mlx-community/Meta-Llama-3.1-8B-Instruct-abliterated-8bit
Text Generation
• 2B • Updated • 87
mlx-community/Meta-Llama-3.1-8B-Instruct-abliterated-2bit
Text Generation
• 0.8B • Updated • 40
mlx-community/Llama-3.1-70B-Japanese-Instruct-2407-4bit
Text Generation
• 11B • Updated • 18
ChrisCoffey/Qwen2-0.5B-test
Text Generation
• 0.5B • Updated • 11
sosoai/hansoldeco-llama-3.1-8b-it-FFT-v0.2-mlx
1B • Updated • 4