Inference Providers
Active filters: chat
RedHatAI/Qwen2.5-14B-Instruct-quantized.w8a8
Text Generation
• 15B • Updated • 102
imkebe/QwQ-32B-Preview-rk3588-1.1.3
Text Generation
• Updated • 6
Text Generation
• Updated Mahmoud-Selim/Llama-Instruct
Text Generation
• 1.7M • Updated • 7
netnk9151/Llama-DNA-1.0-8B-Instruct-Q8_0-GGUF
Text Generation
• 8B • Updated • 10
mradermacher/Llama-DNA-1.0-8B-Instruct-GGUF
8B • Updated • 92
mradermacher/Llama-DNA-1.0-8B-Instruct-i1-GGUF
8B • Updated • 609
sethut/QwQ-32B-Preview-Q8_0-GGUF
33B • Updated • 4
QuantFactory/Llama-DNA-1.0-8B-Instruct-GGUF
Text Generation
• 8B • Updated • 79
• 2
madroid/Qwen2.5-3B-Instruct-4bit-mlx
Text Generation
• 0.5B • Updated • 18
NousResearch/Hermes-3-Llama-3.2-3B-GGUF
3B • Updated • 4.27k
• 73
mlx-community/Josiefied-Qwen2.5-14B-Instruct-abliterated-v4-4-bit
Text Generation
• 2B • Updated • 59
• 1
eligapris/Qwen2.5-Coder-32B-Instruct-Q4_K_M-GGUF
Text Generation
• 33B • Updated • 9
tensorblock/Sailor-0.5B-Chat-GGUF
0.6B • Updated • 78
mlx-community/QwQ-32B-Coder-Fusion-9010-4bit
Text Generation
• 5B • Updated • 23
• 1
tensorblock/Experiment31-7B-GGUF
Text Generation
• 7B • Updated • 8
chende2024/QwQ-32B-Preview-Q4_0-GGUF
33B • Updated • 4
• 1
chende2024/Qwen2.5-1.5B-Instruct-Q4_K_M-GGUF
Text Generation
• 2B • Updated • 24
lianghsun/Llama-3.2-Taiwan-1B-Instruct
Text Generation
• 1B • Updated • 3
zai-org/VisionReward-Image
Text Generation
• Updated • 11
Audio-Text-to-Text
• 0.6B • Updated • 485
• 289
bartowski/Hermes-3-Llama-3.2-3B-GGUF
Text Generation
• 3B • Updated • 14.3k
• 15
ggml-org/Qwen2.5-Coder-1.5B-32B-speculative-GGUF
Text Generation
• 2B • Updated • 84
• 5
mlx-community/Hermes-3-Llama-3.2-3B-4bit
Text Generation
• 0.5B • Updated • 215
• 1
mlx-community/Hermes-3-Llama-3.2-3B-8bit
Text Generation
• 0.9B • Updated • 72
• 1
mlx-community/Hermes-3-Llama-3.2-3B-bf16
Text Generation
• 3B • Updated • 29
mradermacher/Holland-4B-V1-GGUF
5B • Updated • 114
• 1
JackeyLai/Qwen2.5-3B-Instruct-Q4_0-GGUF
Text Generation
• 3B • Updated • 21
JackeyLai/Qwen2.5-7B-Instruct-Q4_0-GGUF
Text Generation
• 8B • Updated • 24
cphan-intersystems/Qwen2.5-Coder-32B-Instruct-Q4_K_M-GGUF
Text Generation
• 33B • Updated • 7