Inference Providers
Active filters: vicuna
digitalpipelines/llama2_13b_chat_uncensored-GGML
eachadea/legacy-ggml-vicuna-13b-4bit
Text Generation
• Updated • 1
• 205
eachadea/legacy-vicuna-13b
Text Generation
• Updated • 10
• 93
Text Generation
• Updated • 18
• 45
Text Generation
• Updated • 23
eachadea/legacy-ggml-vicuna-7b-4bit
Text Generation
• Updated • 79
Text Generation
• Updated • 7
CRD716/ggml-vicuna-1.1-quantized
Text Generation
• Updated • 43
Thireus/Vicuna13B-v1.1-8bit-128g
Text Generation
• Updated • 27
• 16
Text Generation
• Updated • 7
• 28
Mozzipa/ko_vicuna_7b_ggml_q4
Text Generation
• Updated • 8
• 5
Text Generation
• Updated • 12
• 23
Text Generation
• Updated • 10
• 33
Text Generation
• Updated • 43
• 28
Text Generation
• Updated • 55
• 6
Text Generation
• 33B • Updated • 101
• 119
CalderaAI/30B-Lazarus-GGMLv5_1
luffycodes/tutorbot-spock-bio-llama-diff
Text Generation
• Updated • 13
• 2
Text Generation
• Updated • 50
• 7
Text Generation
• Updated • 138
• 11
asedmammad/Vicuna-7B-vanilla-1.1-GGML
Text Generation
• 13B • Updated • 60
• 9
CalderaAI/13B-Ouroboros-GPTQ4bit-128g-CUDA
Text Generation
• Updated • 7
TheBloke/30B-Epsilon-GPTQ
Text Generation
• 33B • Updated • 72
• 6
TheBloke/30B-Epsilon-GGML
Updated • 12
• 9
TheBloke/13B-Ouroboros-GGML
Text Generation
• Updated • 12
• 4
TheBloke/13B-Ouroboros-GPTQ
Text Generation
• 13B • Updated • 26
• 4
TheBloke/13B-BlueMethod-GPTQ
Text Generation
• 13B • Updated • 27
• 6
TheBloke/13B-BlueMethod-GGML
Updated • 6
• 10