Inference Providers
Active filters: 4-bit
iproskurina/bloom-1b7-GPTQ-4bit-g128
Text Generation
• 1B • Updated • 11
Narya-ai/zephyr-7b-sft-lora
Text Generation
• Updated • 4
Jack2022/opt-125m-gptq-4bit
Text Generation
• 0.1B • Updated • 5
Weni/ZeroShot-3.0.1-Mistral-7b-Multilanguage-3.0.3-AWQ
Text Generation
• 7B • Updated • 9
TheBloke/RpBird-Yi-34B-200k-GPTQ
Text Generation
• 34B • Updated • 5
• 3
TheBloke/RpBird-Yi-34B-200k-AWQ
Text Generation
• 34B • Updated • 5
• 1
Abe13/zephyr-7b-sft-lora-1
Text Generation
• Updated • 7
AshanGimhana/quant-model-GPTQ-V1.1
Text Generation
• 0.8B • Updated • 6
Text Generation
• 35B • Updated • 11
• 3
Text Generation
• 35B • Updated • 10
• 2
Abe13/zephyr-7b-sft-lora-2
Text Generation
• Updated • 6
ramgpt/ramgpt-13b-awq-gemm
Question Answering
• 13B • Updated • 1
TheBloke/OpenOrca-Zephyr-7B-GPTQ
Text Generation
• 7B • Updated • 4
• 1
TheBloke/OpenOrca-Zephyr-7B-AWQ
Text Generation
• 7B • Updated • 6
• 1
ramgpt/deepseek-coder-6.7B-GPTQ
Text Generation
• Updated • 11
hnguyen2k/Asclepius-Llama2-7B-GPTQ
Text Generation
• Updated • 4
girrajjangid/Llama-7B-SFT-AWQ
Text Generation
• 7B • Updated • 5
Medilora/guideline-medqa-adapter
Text Generation
• Updated • 6
7B • Updated • 5
heka-ai/llama-2-7b-eytan-test-GPTQ
Text Generation
• 7B • Updated • 6
Text Generation
• 7B • Updated • 6
• 1
Text Generation
• 7B • Updated • 11
• 2
TheBloke/SUS-Chat-34B-AWQ
Text Generation
• 34B • Updated • 12
• 9
TheBloke/SUS-Chat-34B-GPTQ
Text Generation
• 34B • Updated • 10
• 3
TheBloke/NeuralOrca-7B-v1-GPTQ
Text Generation
• 7B • Updated • 4
• 1
TheBloke/NeuralOrca-7B-v1-AWQ
Text Generation
• 7B • Updated • 8
• 1
TheBloke/una-cybertron-7B-v2-AWQ
Text Generation
• 7B • Updated • 14
• 4
TheBloke/una-cybertron-7B-v2-GPTQ
Text Generation
• 7B • Updated • 9