Inference Providers
Active filters: mlx
mlx-community/Llama-3-8b-64k-PoSE-4bit
Text Generation
• 1B • Updated • 13
GreenBitAI/Phi-3-mini-4k-instruct-layer-mix-bpw-2.2-mlx
0.5B • Updated • 9
mlx-community/Llama-3-8b-64k-PoSE-8bit
Text Generation
• 2B • Updated • 15
• 1
mlx-community/Meta-Llama-3-8B-Instruct-64k-4bit
Text Generation
• 1B • Updated • 33
• 1
GreenBitAI/Phi-3-mini-128k-instruct-layer-mix-bpw-3.0-mlx
0.6B • Updated • 7
GreenBitAI/Phi-3-mini-128k-instruct-layer-mix-bpw-2.5-mlx
0.6B • Updated • 5
GreenBitAI/Phi-3-mini-128k-instruct-layer-mix-bpw-2.2-mlx
0.5B • Updated • 6
GreenBitAI/Phi-3-mini-4k-instruct-layer-mix-bpw-3.0-mlx
0.6B • Updated • 4
GreenBitAI/Phi-3-mini-4k-instruct-layer-mix-bpw-2.5-mlx
0.6B • Updated • 4
mlx-community/Llama-3-8B-Instruct-262k-4bit
Text Generation
• 1B • Updated • 18
• 3
mlx-community/Llama-3-8B-Instruct-262k-8bit
Text Generation
• 2B • Updated • 17
• 3
mlx-community/Llama-3-8B-Instruct-262k-unquantized
Text Generation
• 8B • Updated • 11
• 1
mlx-community/Llama-3-Aplite-Instruct-4x8B-MoE-4bit
Text Generation
• 4B • Updated • 24
Text Generation
• 9B • Updated • 10
mradermacher/Llama-3-8B-Instruct-262k-unquantized-GGUF
8B • Updated • 104
• 1
mradermacher/Llama-3-8B-Instruct-262k-unquantized-i1-GGUF
8B • Updated • 251
mayflowergmbh/Llama3_DiscoLM_German_8b_v0.1_experimental-4bit
Text Generation
• 2B • Updated • 10
mayflowergmbh/Llama-3-SauerkrautLM-8b-Instruct-4bit
2B • Updated • 28
• 2
mlx-community/Qwen1.5-110B-4bit
Text Generation
• 17B • Updated • 11
• 1
mlx-community/Qwen1.5-110B-8bit
Text Generation
• 31B • Updated • 11
mlx-community/Qwen1.5-110B-Chat-4bit
Text Generation
• 17B • Updated • 41
• 5
mlx-community/Qwen1.5-110B-Chat-8bit
Text Generation
• 31B • Updated • 15
• 1
mlx-community/nanoLLaVA-4bit
0.7B • Updated • 11
mlx-community/Meta-Llama-3-8B-Instruct
Text Generation
• 8B • Updated • 24
• 2
mlx-community/Swallow-70b-instruct-v0.1-4bit
Text Generation
• 11B • Updated • 15
• 1
mlx-community/UTENA-7B-NSFW-V2-4bit
1B • Updated • 240
• 1
mlx-community/MXLewd-L2-20B-4bit
1B • Updated • 21
mlx-community/Swallow-70b-instruct-v0.1-8bit
Text Generation
• 20B • Updated • 10
mlx-community/Swallow-7b-instruct-v0.1-8bit
Text Generation
• 2B • Updated • 7
mlx-community/Swallow-13b-instruct-v0.1-8bit
Text Generation
• 4B • Updated • 5