Inference Providers
Active filters: meta
mradermacher/llama-3-8b-1m-PoSE-i1-GGUF
8B • Updated • 51
Hanqix/Meta-Llama-3-8B-Instruct-Q3_K_L-GGUF
Text Generation
• 8B • Updated • 5
jenhseb/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 7
Text Generation
• 1B • Updated • 10
• mradermacher/Llama-3-8B-Instruct-64k-GGUF
8B • Updated • 54
• 1
mradermacher/Llama-3-8B-Instruct-64k-i1-GGUF
8B • Updated • 255
• 1
jsbaicenter/Llama-3.2-3b-Instruct-AWQ-4bit-GEMM
Text Generation
• 3B • Updated • 9
skarmani/Llama-3.1-8B-Instruct-Q5_K_M-GGUF
Text Generation
• 8B • Updated • 2
yhkim96/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 2
yhkim96/Llama-3.2-3B-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 10
abdulmannan-01/llama3.1-8b-full-finetuned-rag-alpaca-data
Text Generation
• 8B • Updated • 4
VinayHajare/Llama-3.2-1B-Instruct-Q4_K_M-GGUF
Text Generation
• 1B • Updated • 17
mradermacher/Llama-3-8B-Instruct-v0.4-GGUF
8B • Updated • 65
ddsbg/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 8
tensorblock/Llama-3.3-70B-Instruct-abliterated-GGUF
71B • Updated • 34
• 2
mradermacher/Llama-3-8B-Instruct-v0.4-i1-GGUF
8B • Updated • 182
mradermacher/Guru-Llama-3-8B-GGUF
8B • Updated • 67
tensorblock/Llama3.2-1B-Instruct-GGUF
Text Generation
• 1B • Updated • 27
Text Generation
• Updated • 5
chg0901/llama3_2_3B_Instruct_alphaca_epoch_1
Text Generation
• 3B • Updated • 11
kjaedhrbgo/Meta-Llama-3.1-8B-Q8_0-GGUF
Text Generation
• 8B • Updated • 9
seungwon12/Llama-2-7b-hf-Q4_0-GGUF
Text Generation
• 7B • Updated • 7
seungwon12/Llama-2-7b-hf-Q4_K_M-GGUF
Text Generation
• 7B • Updated • 7
cmcmaster/Llama-3.2-3B-Q4-mlx
Text Generation
• 0.5B • Updated • 20
matrixportalx/Meta-Llama-3-8B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 13
• 1
matrixportalx/Meta-Llama-3-8B-Instruct-Q4_0-GGUF
Text Generation
• 8B • Updated • 13
mlx-community/Llama-3.2-1B-Instruct-MLXTuned
Text Generation
• 1B • Updated • 87
• 5
mradermacher/Trillama-8B-GGUF
8B • Updated • 20
RedHatAI/Llama-3.3-70B-Instruct-quantized.w4a16
Text Generation
• 71B • Updated • 11.2k
• 4
tensorblock/llama-3-8b-instruct-262k-chinese-GGUF
Text Generation
• 8B • Updated • 45