Inference Providers
Active filters: 4-bit
jamesdborin/llama2-13b-chat-4bit-AWQ
Text Generation
• 13B • Updated • 5
• 1
jamesdborin/llama2-70b-chat-4bit-AWQ
Text Generation
• 69B • Updated • 5
• 1
jamesdborin/llama2-7b-base-4bit-AWQ
Text Generation
• 7B • Updated • 7
jamesdborin/llama2-13b-base-4bit-AWQ
Text Generation
• 13B • Updated • 8
jamesdborin/llama2-70b-base-4bit-AWQ
Text Generation
• 69B • Updated • 8
Chat-Error/Rose-Kimiko-20B
Updated • 7
• 2
Weni/ZeroShot-3.1.1-Mistral-7b-Multilanguage-3.0.3-AWQ
Text Generation
• 7B • Updated • 4
TheBloke/OpenHermes-2.5-neural-chat-v3-3-Slerp-GPTQ
Text Generation
• 7B • Updated • 7
• 3
TheBloke/OpenHermes-2.5-neural-chat-v3-3-Slerp-AWQ
Text Generation
• 7B • Updated • 6
• 1
TheBlokeAI/Mixtral-tiny-GPTQ
Text Generation
• 0.2B • Updated • 201
• 3
TheBloke/leo-hessianai-70B-AWQ
Text Generation
• 69B • Updated • 6
TheBloke/leo-hessianai-70B-GPTQ
Text Generation
• 69B • Updated • 92
bienpr/Llama-2-7B-Chat-GPTQ
Text Generation
• 7B • Updated • 5
Chuanming/peft-exercise-1
TheBloke/Mixtral-8x7B-v0.1-GPTQ
Text Generation
• 47B • Updated • 350
• 126
minhhiepcr/quantizated_Mistral-7B-Instruct-v0.1-awq
Text Generation
• 7B • Updated • 6
iproskurina/opt-350m-GPTQ-4bit-g128
Text Generation
• 95.6M • Updated • 13
iproskurina/opt-1.3b-GPTQ-4bit-g128
Text Generation
• 0.4B • Updated • 13
iproskurina/opt-2.7b-GPTQ-4bit-g128
Text Generation
• 0.6B • Updated • 13
iproskurina/opt-6.7b-GPTQ-4bit-g128
Text Generation
• 1B • Updated • 48
ybelkada/llava-1.5-7b-hf-awq
Image-Text-to-Text
• 7B • Updated • 385
• 2
TheBloke/LlamaGuard-7B-AWQ
Text Generation
• 7B • Updated • 111
• 4
marcsun13/Mixtral-tiny-GPTQ
Text Generation
• 0.2B • Updated • 9
TheBloke/Mixtral-8x7B-Instruct-v0.1-GPTQ
Text Generation
• 47B • Updated • 236
• 141
TheBloke/LlamaGuard-7B-GPTQ
Text Generation
• 7B • Updated • 18
• 2
marcsun13/Mixtral-8x7B-v0.1-GPTQ
Text Generation
• 47B • Updated • 11
HydraIndicLM/mistral-MoQlora-punjabi-expert
Text Generation
• Updated • 4
HydraIndicLM/mistral-MoQlora-tamil-expert
Text Generation
• Updated • 4
TheBloke/mixtral-8x7b-v0.1-AWQ
Text Generation
• 47B • Updated • 211
• 11
TheBloke/Mixtral-8x7B-Instruct-v0.1-AWQ
Text Generation
• 47B • Updated • 1.05k
• 59