RedHatAI/Meta-Llama-3.1-70B-Instruct-quantized.w4a16 Text Generation • 71B • Updated Feb 12, 2025 • 8.78k • 33
RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w4a16 Text Generation • 8B • Updated Jul 10 • 50.1k • 30
RedHatAI/Meta-Llama-3.1-405B-Instruct-quantized.w4a16 Text Generation • 406B • Updated Aug 19 • 313 • 11
RedHatAI/Meta-Llama-3.1-405B-Instruct-FP8-dynamic Text Generation • 406B • Updated Aug 19 • 1.98k • 15
RedHatAI/Meta-Llama-3.1-405B-Instruct-quantized.w8a8 Text Generation • 406B • Updated Aug 19 • 70 • 2
RedHatAI/Meta-Llama-3.1-70B-Instruct-quantized.w8a8 Text Generation • 71B • Updated Aug 19 • 1.01k • 21
RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w8a8 Text Generation • 8B • Updated Aug 19 • 12.4k • 20
RedHatAI/Meta-Llama-3.1-70B-Instruct-quantized.w8a16 Text Generation • 71B • Updated Aug 19 • 124 • 5
RedHatAI/Meta-Llama-3.1-405B-Instruct-quantized.w8a16 Text Generation • 406B • Updated Aug 19 • 108 • 2