File size: 584 Bytes
a671b16
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
Qwen3.6-27B NVIDIA NVFP4 no-MTP 16GB deployment derivative
Copyright 2026 Wilson Zhang (scripts and original documentation only)

This distribution contains documentation and tooling for a modified derivative of:
- Qwen/Qwen3.6-27B
- nvidia/Qwen3.6-27B-NVFP4
- utautako/Qwen3.6-27B-NVIDIA-NVFP4-MTP-GGUF

Modification: the single embedded MTP / next-token-prediction layer was physically removed; retained main-model tensors were copied without requantization.

No affiliation with or endorsement by Qwen, Alibaba, NVIDIA, utautako, Hugging Face, or the llama.cpp project is implied.