MTP support? Why no Vision?

#3
by unoid - opened

base qwen models had vision already for you to reuse, also MTP, did you retrain the MTP layers?

If not, why not include them, vision and MTP are free lunch

Thanks for your thoughtful questions!
Our optimization for KAT-Coder-V2.5-Dev is focused purely on agentic coding scenarios. This checkpoint targets text-only code agent tasks, and we have not conducted any training for vision capabilities. Therefore, we do not recommend using this model for multimodal vision-related workloads. If you require vision support, the original Qwen3.6-35BA3B base model would be the better choice.

Sign up or log in to comment