Better than the original, even at coding tasks.

#10
by sasih - opened

Honestly it has been night and day for me comparing Qwen3.6-27B-UD-Q4_K_XL.gguf to this heretic MTP version. With identical settings on my llama.cpp server I get way better coding-agent results with this abliterated-MTP model than I do the original at the same bit depth. Thank you!

Owner
β€’
edited May 23

Honestly it has been night and day for me comparing Qwen3.6-27B-UD-Q4_K_XL.gguf to this heretic MTP version. With identical settings on my llama.cpp server I get way better coding-agent results with this abliterated-MTP model than I do the original at the same bit depth. Thank you!

Yes, the original model that Unsloth is distributing very censored (92/100 refusals), while my version right here has 6/100 refusals with a KL divergence of 0.0021, that's basically the original model without the safeguards which are constantly holding back the model, basically a censored model with safeguards has to be constantly wasting some of it's capabilities on evaluating if a prompt is "okay" and "safe" to proceed with before consenting or refusing to proceed, basically the model is being held by a tight leash by the safeguards and it can not do the job as well as it could because it is constantly holding back by the safeguwards and this is having a negative effect on the model's capabilities, an uncensored model doesn't have that and instead just take a look at the prompt and just proceed with the workflow.

Sign up or log in to comment