Inference Providers
Active filters: 8-bit
mNLP-project/gpt2-safety-GPTQ-8bit
Text Generation
• 0.2B • Updated • 3
Text Generation
• 0.1B • Updated • 5
cs552-mlp/phi3-gptq-8bits
Text Generation
• 4B • Updated • 4
Text Generation
• 0.6B • Updated • 8
grimjim/Llama-3-Luminurse-v0.2-OAS-8B-8bpw-exl2
Text Generation
• Updated • 5
PeterAM4/EPFL-TA-Meister-8bit
Text Generation
• 8B • Updated • 7
cs552-mlp/phi3-lora-gptq-8bits
Text Generation
• 4B • Updated • 6
cs552-mlp/phi3-lora-sciq3-gptq
Text Generation
• 4B • Updated • 4
Shaleen123/llama3-webinstruct-8bit
Text Generation
• 8B • Updated • 5
mNLP-project/baseline-gpt2-quantized
Text Generation
• 0.2B • Updated • 5
bilalkakar/BLOOM560-MCQ-Quantized
Text Generation
• 0.6B • Updated • 5
Veture/merged_autoGPTQ_dpo
Text Generation
• 2B • Updated • 5
nm-testing/tinyllama-oneshot-w8w8-test-static-shape-change
Text Generation
• 1B • Updated • 65k
MoslemTCM/Mistral-7B-finetuned-vigoo-with-lora
Text Generation
• 7B • Updated • 7
• 1
SotiriosKastanas/microsoft-Orca-2-7b-8-bit-gptq
Text Generation
• 7B • Updated • 5
SotiriosKastanas/mistralai-Mistral-7B-v0.1-8-bit-gptq
Text Generation
• 7B • Updated • 8
SotiriosKastanas/berkeley-nest-Starling-LM-7B-alpha-8-bit-gptq
Text Generation
• 7B • Updated • 6
SotiriosKastanas/Intel-neural-chat-7b-v3-8-bit-gptq
Text Generation
• 7B • Updated • 4
SotiriosKastanas/meta-llama-Meta-Llama-3-8B-8-bit-gptq
Text Generation
• 8B • Updated • 6
SotiriosKastanas/microsoft-Orca-2-13b-8-bit-gptq
Text Generation
• 13B • Updated • 5
SotiriosKastanas/microsoft-Phi-3-medium-4k-instruct-8-bit-gptq
Text Generation
• 14B • Updated • 6
• 2
thewordsmiths/Llama_SciQ_8bits
Text Generation
• 8B • Updated • 8
Statuo/EtherealRainbow-v0.2-8B_EXL2_8BPW
Text Generation
• Updated • 3
• 3
SotiriosKastanas/google-gemma-7b-8-bit-gptq
Text Generation
• 9B • Updated • 5
njaana/phi3-mini-new-model-demo2
Text Generation
• 2B • Updated • 5
Doub7e/llama-3-8b-dpo-distilabel-epfl-sft-bnb8bit
Text Generation
• 8B • Updated • 5
SotiriosKastanas/lmsys-vicuna-7b-v1.5-8-bit-gptq
Text Generation
• 7B • Updated • 7
SotiriosKastanas/lmsys-vicuna-13b-v1.5-8-bit-gptq
Text Generation
• 13B • Updated • 6
Text Generation
• 0.8B • Updated • 7
Doub7e/llama-3-8b-dpo-distilabel-epfl-sft-gptq8bit
Text Generation
• 8B • Updated • 4