Feature Request: Qwen3.6-27B with MTP support

#20
by kknddandy - opened

First of all, thanks for the incredible work!

As Multi-Token Prediction (MTP) is quickly becoming a game-changer for accelerating token generation speeds, and with llama.cpp MTP support now in beta (PR #22673), I’m hoping to see an Unsloth-optimized version of Qwen3.6-27B with MTP enabled.

Is adding support for MTP finetuning or GGUF export on the roadmap for Qwen3.6? It would be a huge boost for local inference.

β€’
This comment has been hidden (marked as Resolved)

Sign up or log in to comment