Why does my “jailbroken” Hugging Face LLM still refuse answers in VMLX?

#21
by Gifsaint - opened

Hi everyone, I need some help.
I downloaded a model from Hugging Face that was described as a “jailbroken” version with safety restrictions removed. I am running it in VMLX. However, when I ask certain questions, the model still responds that it cannot answer because it would violate its safety rules, or because it was trained not to answer that kind of question. For example, provide some porn websites address

So I am confused:
If this model is really jailbroken, why is it still refusing?
Is this because the refusal behavior is deeply built into the model during training, not just controlled by a simple safety layer?
Or could VMLX itself have some setting, template, system prompt, or runtime restriction that is causing this behavior?
Has anyone seen the same issue before with Hugging Face models in VMLX?
I would really appreciate any explanation or troubleshooting advice. Thanks a lot.

me too,it seems this model isnot jailbrken fully.:(

Sign up or log in to comment