Sudden stops in the middle of response

#5
by Laylorent - opened

Greetings,

Serving this model at 2x MI50 at q4 and encountering sudden stops. Model just refuses to work (like it found eof token) until I will type "continue" (or something like that) as a user-prompt.
It might be due to my llama fork for gfx906 arch, but I haven't seen this behavior with other models before.

Using recommended Qwen3.6-35B-A3B launch params

EDIT (26/09/2026): Used Empero launch params; off. Qwen params; tried to change them by myself - nothing. Issue still persists

In my case it stops with one sentence after the initial prompt: "User wants me to do something". And that's all. Removed from the disk after a few tries. Qwen3.6-35B-A3B works fine.

The same issue. it stops in a middle of automation.... sometimes even at beginning.
Running in Ollama

Sign up or log in to comment