Inference Providers
Active filters: 4-bit
Text Generation
• 0.1B • Updated • 6
TheBloke/Autolycus-Mistral_7B-AWQ
Text Generation
• 7B • Updated • 8
• 1
TheBloke/Autolycus-Mistral_7B-GPTQ
Text Generation
• 7B • Updated • 6
• 1
TheBloke/Free_Sydney_V2_Mistral_7b-AWQ
Text Generation
• 7B • Updated • 6
• 1
TheBloke/Generate_Question_Mistral_7B-AWQ
Text Generation
• 7B • Updated • 10
• 2
TheBloke/Free_Sydney_V2_Mistral_7b-GPTQ
Text Generation
• 7B • Updated • 5
• 1
TheBloke/Writing_Partner_Mistral_7B-AWQ
Text Generation
• 7B • Updated • 14
• 1
TheBloke/Generate_Question_Mistral_7B-GPTQ
Text Generation
• 7B • Updated • 9
• 3
TheBloke/Writing_Partner_Mistral_7B-GPTQ
Text Generation
• 7B • Updated • 6
TheBloke/llama2_7b_merge_orcafamily-AWQ
Text Generation
• 7B • Updated • 7
TheBloke/llama2_7b_merge_orcafamily-GPTQ
Text Generation
• 7B • Updated • 4
TheBloke/mistral_7b_norobots-GPTQ
7B • Updated TheBloke/mistral_7b_norobots-AWQ
7B • Updated TheBloke/zephyr_7b_norobots-AWQ
7B • Updated • 1
TheBloke/zephyr_7b_norobots-GPTQ
7B • Updated • 1
gustavovm/neural-chat-7b-pt-br-quant
Text Generation
• Updated • 5
TheBloke/Yarn-Llama-2-70B-32k-AWQ
Text Generation
• 69B • Updated • 9
• 2
TheBloke/Yarn-Llama-2-70B-32k-GPTQ
Text Generation
• 69B • Updated • 6
LoftQ/Llama-2-7b-hf-4bit-64rank
Text Generation
• 7B • Updated • 44
• 1
Text Generation
• 13B • Updated • 58
• 27
Text Generation
• 7B • Updated • 86
• 1
Text Generation
• 13B • Updated • 24
• 7
Text Generation
• 7B • Updated • 10
• 4
TheBloke/Qwen-7B-Chat-GPTQ
Text Generation
• 8B • Updated • 77
• 3
TheBloke/Qwen-7B-Chat-AWQ
Text Generation
• 8B • Updated • 59
• 9
TheBloke/Qwen-14B-Chat-AWQ
Text Generation
• 14B • Updated • 61
• 11
marcchew/TinyLLaMA-1.1B-OrcaPlatty-GPTQ-4bit
Text Generation
• 1B • Updated • 7
Eichhof/Llama-2-13b-chat-hf-Einstein-GPTQ-C4
Text Generation
• 13B • Updated Text Generation
• Updated • 6
BradClem/chess-chat-1.0-3b-Q4
Text Generation
• 3B • Updated • 6