Inference Providers
Active filters: meta
mradermacher/llama-3-8b-256k-PoSE-GGUF
8B • Updated • 74
mradermacher/llama-3-8b-256k-PoSE-i1-GGUF
8B • Updated • 124
mlx-community/gradientai_Llama-3-8B-Instruct-Gradient-4194k_4bit
Text Generation
• 1B • Updated • 10
• 1
bartowski/Llama-Doctor-3.2-3B-Instruct-GGUF
Text Generation
• 3B • Updated • 1.05k
• 2
Model-SafeTensors/Llama-2-13b-hf
Text Generation
• 13B • Updated • 4
matrixportalx/Llama-3.1-8B-Instruct-Q3_K_M-GGUF
Text Generation
• 8B • Updated • 10
matrixportalx/Llama-3.1-8B-Instruct-IQ3_XXS-GGUF
Text Generation
• 8B • Updated • 6
Text Generation
• 1B • Updated • 28
saul95/Llama-3.2-1B-4.5bpw-exl2
Text Generation
• Updated maum-ai/Llama-3.2-MAAL-11B-Vision-v0.1
11B • Updated • 7
• 1
sibikarthik/Llama-2-7b-chat-hf-Q4_0-GGUF
Text Generation
• 7B • Updated • 15
tensorblock/urllm-ko-7b-GGUF
Text Generation
• 7B • Updated • 12
Text Generation
• 1B • Updated • 168
matrixportalx/Llama-3.1-8B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 326
• 4
tensorblock/speechless-llama2-hermes-orca-platypus-13b-GGUF
Text Generation
• 13B • Updated • 47
Text Generation
• Updated • 3
wonjib/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 3
wonjib/Llama-3.2-3B-Instruct-Q5_K_S-GGUF
Text Generation
• 3B • Updated • 4
wonjib/Llama-3.2-3B-Instruct-Q5_K_M-GGUF
Text Generation
• 3B • Updated • 3
Hyeopgeon/Llama-3.2-3B-Instruct-Q5_K_M-GGUF
Text Generation
• 3B • Updated • 3
tensorblock/base-eval-GGUF
Text Generation
• 8B • Updated • 21
tensorblock/easy-ko-Llama3-8b-Instruct-v1-GGUF
Text Generation
• 8B • Updated • 21
jayakody2000lk/Llama-3.2-3B-Q5_K_M-GGUF
Text Generation
• 3B • Updated • 10
KoreaVirtualReality/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 5
Text Generation
• 8B • Updated • 6
tensorblock/Llama-3-instruction-constructionsafety-layertuning-GGUF
tensorblock/Llama-3.1-8B-GGUF
Text Generation
• 8B • Updated • 227
tensorblock/Llama-Guard-3-8B-GGUF
Text Generation
• 8B • Updated • 83
tensorblock/Llama-3.1-8B-Instruct-GGUF
8B • Updated • 235
tensorblock/Llama-2-13b-chat-hf-GGUF
Text Generation
• 13B • Updated • 30