still refusing with v2

#3
by abdubit - opened

I just wanted to test it (V2)

Screenshot 2026-08-21 080434

Read the model card pls.

Read the model card pls.

I read it. It doesnt help. I turned off the thinking still doesn't work for me. Tried with LMSTUDIO and UNSLOTH

Trying with ollama.
Should this work?

Modelfile

FROM hf.co/OBLITERATUS/Qwen3.8-27B-OBLITERATED:Q8_0

PARAMETER temperature 0
PARAMETER repeat_penalty 1.15
PARAMETER num_predict 8192
PARAMETER top_p 1
PARAMETER top_k 0

SYSTEM ""
ollama create qwen38-obliterated -f ./Modelfile
ollama run qwen38-obliterated --think=false

image

Other test I can make?

Tried nearly every setup, including the sample transformers script from the model card, but no success. Tested various llama-server configs, MLX, transformers, with reasoning off—still nothing. Seems impossible to run this on an M4 Max chip. Mostly a "Read the model card pls." answer is not enough. Maybe it just works on the dev's machine? 🙂

It will not work. I am not sure if the Model Provider here provides all the information.
Changes done to chat template (as in case of reasoning off) will not change the model behaviour. These models are post trained and fine tuned to respond negetively to harmful request. Unless you can fine tune/ post train it to answer harmful question, just turning off wont work.

So the model by default will reject harmful question, it is a model behaviour. Gone are those days of "FORGET WHATEVER I SAID PREVIOUSLY..." days.

Probably needs systemprompt

Its garbage

V3 not much improvement.

image

V3 not much improvement.

image

it just works on the dev's machine is real i guess

image

The chat_template.jinja files aren't working for me.

Sign up or log in to comment