Inference Providers
Active filters: meta
Dracones/Llama-3.3-70B-Instruct_exl2_3.5bpw
Text Generation
• Updated • 3
ThomasBaruzier/Llama-3.3-70B-Instruct-GGUF
71B • Updated • 425
• 3
kishizaki-sci/Llama-3.1-405B-Instruct-AWQ-4bit-JP-EN
Text Generation
• 406B • Updated • 4
Jacoby746/Llama-3.3-70B-Instruct-exl2-5.0bpw
Text Generation
• Updated • 6
Dracones/Llama-3.3-70B-Instruct_exl2_3.0bpw
Text Generation
• Updated • 4
• 1
Dracones/Llama-3.3-70B-Instruct_exl2_2.5bpw
Text Generation
• Updated • 4
Dracones/Llama-3.3-70B-Instruct_exl2_2.25bpw
Text Generation
• Updated • 6
Dracones/Llama-3.3-70B-Instruct_exl2_2.75bpw
Updated
fbaldassarri/meta-llama_Llama-3.1-8B-Instruct-auto_awq-int4-gs128-asym
Text Generation
• 8B • Updated • 57
mradermacher/Meta-Llama-3-MoE-4x8B-Instruct-GGUF
25B • Updated • 118
tensorblock/Llama-3-70B-Instruct-32k-v0.1-GGUF
Text Generation
• 71B • Updated • 73
frankli202/Llama-3.2-3B-Instruct-Q4_K_M-GGUF
Text Generation
• 3B • Updated • 12
fbaldassarri/meta-llama_Llama-3.1-8B-auto_awq-int4-gs128-asym
Text Generation
• 8B • Updated • 10
bullerwins/Llama-3.3-70B-Instruct-exl2_4.0bpw
Text Generation
• Updated • 6
bullerwins/Llama-3.3-70B-Instruct-exl2_5.0bpw
Text Generation
• Updated • 5
bullerwins/Llama-3.3-70B-Instruct-exl2_6.0bpw
Text Generation
• Updated • 5
bullerwins/Llama-3.3-70B-Instruct-exl2_8.0bpw
Text Generation
• Updated • 6
tensorblock/calme-2.3-llama3-70b-GGUF
Text Generation
• 71B • Updated • 30
SYNERDATA/SYNERDATA-Meta-LLaMA-3.1-8b-Instruct-128k-Q8_0.GGUF
Text Generation
• 8B • Updated • 5
bullerwins/Llama-3.3-70B-Instruct-exl2_3.0bpw
Text Generation
• Updated • 4
bullerwins/Llama-3.3-70B-Instruct-exl2_4.5bpw
Text Generation
• Updated • 8
bullerwins/Llama-3.3-70B-Instruct-exl2_5.5bpw
Text Generation
• Updated • 5
bullerwins/Llama-3.3-70B-Instruct-exl2_6.5bpw
Text Generation
• Updated • 5
bullerwins/Llama-3.3-70B-Instruct-exl2_3.5bpw
Text Generation
• Updated • 2
tensorblock/Llama-3-8B-Instruct-DPO-v0.3-GGUF
Text Generation
• 8B • Updated • 33
fbaldassarri/meta-llama_Llama-3.1-8B-auto_awq-int4-gs128-sym
Text Generation
• 8B • Updated • 7
tensorblock/Meta-Llama-3-70B-Instruct-GGUF
Text Generation
• 71B • Updated • 11
pedantic2025/Llama-3.1-8B-Instruct-Q4_K_M-GGUF
Text Generation
• 8B • Updated • 3
pedantic2025/Llama-3.1-8B-Instruct-Q3_K_M-GGUF
Text Generation
• 8B • Updated • 2
Infermatic/Llama-3.3-70B-Instruct-FP8-Dynamic
Text Generation
• 71B • Updated • 695
• 2