We have updated the BananaMind Base Bench leaderboard! We now have these benchmark cards, they make it way easier to see which models are actually good! We've also added the model advisor. It asks you what you want to use the model for and the parameter range and gives you the best model for your task!
I'm officially canceling my Hugging Face Pro subscription today. I supported this platform because it stood for true openness and neutrality. This acquisition by NVIDIA fundamentally changes that.
Hereβs why Iβm against this deal: - Neutrality is dead. NVIDIA is a US-based company. This means US regulations will inevitably dictate platform policies, creating direct pressure on Chinese developers and anyone building open-weight models outside the US. - Community over bureaucracy. NVIDIA is a massive, slow-moving corporation. This acquisition will likely drown the community in corporate processes and commercial interests. Soon, uploading a simple finetune might become a bureaucratic nightmare. - Open vs. Proprietary. Hugging Face was built on open-source ideals. NVIDIA? They are a fiercely proprietary hardware company with a minimal track record of meaningful open-source contributions. They sell chips, not freedom. - And to add insult to injury, NVIDIA has practically abandoned consumer RTX GPUs in 2026 to chase data center profits. Why would I pay them for "openness" when they've turned their back on the very developers who built this ecosystem?
I paid for openness. Not for a corporate takeover.
Qwen 3.5 9B - The Defiant, 27B power ; now with Qwen 3.8 Reasoning modes.
640 ARC-C for both 8bit and 4bit. Model exceeds 7 of 7 benchmarks for Qwen 3.5 9B, Qwen3.5 27B, Qwen3.6 35B-A3B, and meets Qwen 3.6 27B in some cases... and it does so in 4bit and 8bit. Regular and MTP (fast) NEO IMATRIX GGUFs provided. (this model is part of the Qwen 3.6 27B Fable Fusion 711 pipelines: 2200+ likes, 3 million + downloads)
NEW - Qwen 3.8 Reasoning Modes: 2 MTP quants (Q6/Q8) Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level). Model name has "plusIQ" in the name.
(there is also a extra robust "tools" version too.)
176 models, 54 orgs, 5 benchmarks, and a whole community of support!
Thanks to everyone whoβs contributed models, reported issues, suggested benchmark improvements, or used the leaderboard to compare and evaluate small language models.
Itβs been awesome watching the leaderboard grow into a broader community resource for transparent and reproducible SLM evaluation.