Inference Providers
Active filters: chat
mradermacher/Llama-3.3-SuperSwallowX-70B-Instruct-v0.1-GGUF
71B • Updated • 15
Text Generation
• 22B • Updated • 4
• 5
mradermacher/Mistral-portuguese-luana-7b-chat-GGUF
7B • Updated • 148
• 1
gghfez/WizardLM-2-22b-RP-GGUF
Text Generation
• 22B • Updated • 32
• 1
gghfez/WizardLM-2-22b-RP-AWQ
Text Generation
• 22B • Updated • 3
gghfez/WizardLM-2-22B-RP-exl2
Text Generation
• Updated • 9
mradermacher/Mistral-portuguese-luana-7b-chat-i1-GGUF
7B • Updated • 505
• 1
taobao-mnn/Meta-Llama-3-8B-Instruct-MNN
Text Generation
• Updated • 30
tensorblock/QwQ-32B-Coder-Fusion-9010-GGUF
Text Generation
• 33B • Updated • 32
• 2
tensorblock/Llama-3.1-SuperSwallow-70B-Instruct-v0.1-GGUF
Text Generation
• 71B • Updated • 9
tensorblock/Llama-DNA-1.0-8B-Instruct-GGUF
Text Generation
• 8B • Updated • 42
tensorblock/Qwen2.5-Coder-14B-Instruct-abliterated-GGUF
Text Generation
• 15B • Updated • 113
tensorblock/Hermes-3-Llama-3.2-3B-GGUF
3B • Updated • 127
• 2
mradermacher/WizardLM-2-22b-RP-GGUF
22B • Updated • 24
• 3
pipihand01/QwQ-32B-Preview-abliterated-linear75
Text Generation
• 33B • Updated • 4
mradermacher/WizardLM-2-22b-RP-i1-GGUF
22B • Updated • 280
• 4
tensorblock/Qwen2.5-3B-Instruct-abliterated-GGUF
Text Generation
• 3B • Updated • 157
pipihand01/QwQ-32B-Preview-abliterated-linear75-GGUF
33B • Updated • 18
akhbar/Qwen2.5-32B-Instruct-abliterated-8bit-128g-actorder_True-GPTQ
Text Generation
• 33B • Updated • 65
quantflex/SmallThinker-3B-Preview-abliterated-GGUF
Text Generation
• 3B • Updated • 82
• 2
tensorblock/Qwen2.5-Coder-3B-Instruct-GGUF
Text Generation
• 3B • Updated • 124
• 1
taobao-mnn/Llama-2-7b-chat-ms-MNN
Text Generation
• Updated • 17
taobao-mnn/TinyLlama-1.1B-Chat-v1.0-MNN
Text Generation
• Updated • 20
taobao-mnn/Meta-Llama-3.1-8B-Instruct-MNN
Text Generation
• Updated • 42
Text Generation
• 2B • Updated • 9
viethq5/Qwen2.5-0.5B-Instruct-f16
Text Generation
• 0.5B • Updated • 8
rockon1095/Qwen2-7B-Instruct-Q4_0-GGUF
Text Generation
• 8B • Updated • 8
DBMe/Monstral-123B-v2-2.85bpw-h6-exl2
Text Generation
• Updated • 7
• 1
tensorblock/Llama-3.3-70B-Instruct-ablated-GGUF
71B • Updated • 23
Felladrin/gguf-Q4_0-Qwen2.5-Coder-32B-Instruct
Text Generation
• 33B • Updated • 18