Wrong size of UD-Q8_K_XL ?

#4
by johnsmithxx - opened

Non-MTP version of Qwen3.5-122B-A10B UD-Q8_K_XL: 171 GB
MTP version of Qwen3.5-122B-A10B UD-Q8_K_XL: 135 GB

All MTP versions seem to be expectably slightly larger than non-MTP version except this one that's significantly smaller. Is this correct?

I don't think so. I can't get it to stay loaded in LM Studio - likely some tensors aren't what's being expected.

Sign up or log in to comment