view article Article The Fast Gemma Challenge: our verified-SOTA recipe, in full FINAL-Bench • 27 days ago • 24
deepseek-ai/DeepSeek-R1-Distill-Qwen-32B Text Generation • 33B • Updated Feb 24, 2025 • 593k • • 1.6k
meta-llama/Llama-4-Scout-17B-16E-Instruct Image-Text-to-Text • 109B • Updated May 22, 2025 • 259k • • 1.34k
nvidia/Llama-3.1-Nemotron-8B-UltraLong-4M-Instruct Text Generation • 8B • Updated Apr 17, 2025 • 1.31k • • 126
meta-llama/Llama-3.1-8B-Instruct Text Generation • 8B • Updated Sep 25, 2024 • 5.88M • • 6.7k