Thank you so much for tuning this model

#14
by tidjei43 - opened

Thank you so much for tuning this model. Finally, best everyday model for working with literature that easily outperforms Qwen 3.7 Plus, let alone Qwen 3.5 397B. And a huge thank you for the model being MTP-free, because at the moment the speed increase from MTP in llama.cpp is very small for large MoE models, while the drop in prompt processing speed and the memory overhead are far too high for MTP to make sense on CPU+GPU. This is the first fine-tune that has actually made the model better, at least in my daily use.

Sign up or log in to comment