Quantization model context length issue in LM Studio

#3
by boychen321 - opened

The quantized model is losing context information; LM Studio only displays 4096 context values ​​instead of the 262144 supported by the model itself. Hopefully, the authors can fix this issue.

Just to clarify, this issue affects the Qwopus3.5-27B-v3.5-IQ4_XS model.

I saw this issue too, I re-quantized the BF16 version into IQ4_XS on my own machine and it runs fine. I'll upload that version on my account.

Sign up or log in to comment