grug-27b-mtp-gguf

grug rock WITH prediction head inside. Q8_0 / Q5_K_M / Q4_K_M, mtp tensors included for llama.cpp speculative decoding (needs build with qwen3_5 MTP support). draft head trained on grug data: t+2 agreement 89.3%. main card: grug-27b-mtp. no-MTP rocks: grug-27b-gguf. grug made by ProCreations.

Downloads last month
173,967
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for ProCreations/grug-27b-mtp-gguf

Base model

Qwen/Qwen3.6-27B
Quantized
(3)
this model

Collection including ProCreations/grug-27b-mtp-gguf