Inference Providers
Active filters: meta
mradermacher/komt-llama2-7b-v1-i1-GGUF
7B • Updated • 16
• 1
shesshan/dumi-llama3.2-3b-instruct
Text Generation
• 3B • Updated • 5
tensorblock/Meta-Llama-3.1-8B-Instruct-reuploaded-GGUF
tensorblock/Llama-3-Groq-8B-Tool-Use-GGUF
Text Generation
• 8B • Updated • 41
• 1
unsloth/Llama-3.2-11B-Vision-Instruct-unsloth-bnb-4bit
Image-Text-to-Text
• 11B • Updated • 2k
• 29
da-gRu/Meta-Llama-3-8B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 3
gautamgc17/llama3.2-vlm-torchtune
Image-Text-to-Text
• 11B • Updated • 4
NeuraLakeAi/iSA-02-Nano-1B-Preview
1B • Updated • 49
• 5
gurro/Llama-3.1-8B-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 6
• 1
amd/Llama-2-7b-chat-hf-awq-g128-int4-asym-fp16-onnx-dml
Text Generation
• Updated amd/Llama-2-7b-hf-awq-g128-int4-asym-fp16-onnx-hybrid
Text Generation
• Updated • 7
gurro/Llama-3.1-8B-Q4_0-GGUF
Text Generation
• 8B • Updated • 5
amd/Llama-2-7b-chat-hf-awq-g128-int4-asym-fp16-onnx-hybrid
Text Generation
• Updated • 5
serena97/Llama-3.2-3B-Instruct-Q8_0-GGUF
Text Generation
• 3B • Updated • 3
jsjsam98/Llama-3.2-1B-Q8_0-GGUF
Text Generation
• 1B • Updated • 10
• 1
tekiny/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 5
• 1
tekiny/Llama-3.1-8B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 8
• 1
tekiny/Llama-3.2-3B-Instruct-Q5_K_S-GGUF
Text Generation
• 3B • Updated • 1
• 1
tekiny/Llama-3.2-1B-Instruct-Q5_K_S-GGUF
Text Generation
• 1B • Updated • 1
• 1
tekiny/Meta-Llama-3-8B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 1
• 1
gurro/Llama-3.1-8B-IQ4_XS-GGUF
Text Generation
• 8B • Updated • 6
• 1
tensorblock/Llama-3-8B-Instruct-v0.4-GGUF
Text Generation
• 8B • Updated • 58
tensorblock/llama-3-8b-chat-GGUF
Text Generation
• 8B • Updated • 38
tensorblock/Meta-Llama-3-8B-hf-GGUF
Text Generation
• 8B • Updated • 32
fbaldassarri/meta-llama_Llama-3.2-3B-Instruct-auto_awq-int4-gs128-asym
Text Generation
• 4B • Updated • 44
mradermacher/Meta-Llama-3-70B-hf-GGUF
71B • Updated • 19
knownasSohan/llama3.2-vlm-torchtune
Image-Text-to-Text
• 11B • Updated • 4
QuantFactory/Open-Insurance-LLM-Llama3-8B-GGUF
Text Generation
• 8B • Updated • 20
• 6
jwiggerthale/Llama-3.2-3B-Q2_K-GGUF
Text Generation
• 3B • Updated • 11
jwiggerthale/Llama-3.2-3B-Q4_0-GGUF
Text Generation
• 3B • Updated • 10