Inference Providers
Active filters: chat
mradermacher/Qwentile2.5-32B-Instruct-i1-GGUF
33B • Updated • 74
• 3
ModelCloud/QwQ-32B-Preview-gptqmodel-4bit-vortex-v3
Text Generation
• 33B • Updated • 16
• 14
TheBlueObserver/Qwen2.5-32B-Instruct-MLX-8777b
Text Generation
• 7B • Updated • 14
IntelligentEstate/Replicant_Warder-QwenStar-3B-iQ5_K_S
Text Generation
• 3B • Updated • 22
• 1
Text Generation
• 15B • Updated • 5
mlx-community/Josiefied-Qwen2.5-3B-Instruct-abliterated-v1-8-bit
Text Generation
• 0.9B • Updated • 85
mlx-community/Josiefied-Qwen2.5-3B-Instruct-abliterated-v1-4-bit
Text Generation
• 0.5B • Updated • 760
Ronaldus/QwQ-32B-Preview-Q4_K_M-GGUF
33B • Updated • 2
mlx-community/Josiefied-Qwen2.5-3B-Instruct-abliterated-v1
Text Generation
• 3B • Updated • 24
NaniDAO/Llama-3.3-70B-Instruct-ablated
Text Generation
• 71B • Updated • 33
• • 21
Vijay109/Qwen2.5-3B-Instruct-Q8_0-GGUF
Text Generation
• 3B • Updated • 21
TheBlueObserver/Qwen2.5-32B-Instruct-MLX-393a7
Text Generation
• 9B • Updated • 13
itlwas/Qwen2.5-Coder-7B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 31
itlwas/QwQ-32B-Preview-Q4_K_M-GGUF
33B • Updated • 15
itlwas/Qwen2.5-Coder-0.5B-Instruct-Q4_K_M-GGUF
Text Generation
• 0.5B • Updated • 74
itlwas/Qwen2.5-Coder-1.5B-Instruct-Q4_K_M-GGUF
Text Generation
• 2B • Updated • 49
itlwas/Qwen2.5-Coder-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 184
itlwas/Qwen2.5-Coder-14B-Instruct-Q4_K_M-GGUF
Text Generation
• 15B • Updated • 502
• 1
itlwas/Qwen2.5-Coder-32B-Instruct-Q4_K_M-GGUF
Text Generation
• 33B • Updated • 43
tensorblock/llama-2-7b-chat-GGUF
mradermacher/Llama-3.3-70B-Instruct-ablated-GGUF
71B • Updated • 131
mradermacher/Llama-3.3-70B-Instruct-ablated-i1-GGUF
71B • Updated • 193
• 1
bartowski/Llama-3.3-70B-Instruct-ablated-GGUF
Text Generation
• 71B • Updated • 1.55k
• 16
typicalUnidos/Qwen2.5-7B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 12
Delta-Vector/Control-Nanuq-8B-exl2
QuantFactory/Multilingual-SaigaSuzume-8B-GGUF
8B • Updated • 68
• 2
ModelCloud/Falcon3-10B-Instruct-gptqmodel-4bit-vortex-v1
Text Generation
• 10B • Updated • 7
• 3
mradermacher/Qwen2.5-14B-Instruct-abliterated-v2-GGUF
15B • Updated • 6.81k
• 9
itlwas/Sailor-0.5B-Chat-Q4_K_M-GGUF
0.6B • Updated • 8
itlwas/Sailor-1.8B-Chat-Q4_K_M-GGUF
2B • Updated • 12