Qwen3.6-27B NVIDIA NVFP4 no-MTP 16GB deployment derivative Copyright 2026 Wilson Zhang (scripts and original documentation only) This distribution contains documentation and tooling for a modified derivative of: - Qwen/Qwen3.6-27B - nvidia/Qwen3.6-27B-NVFP4 - utautako/Qwen3.6-27B-NVIDIA-NVFP4-MTP-GGUF Modification: the single embedded MTP / next-token-prediction layer was physically removed; retained main-model tensors were copied without requantization. No affiliation with or endorsement by Qwen, Alibaba, NVIDIA, utautako, Hugging Face, or the llama.cpp project is implied.