azukivc's picture jenerallee78's picture
Duplicate from jenerallee78/Qwen3.8-27B-Abliterated-SFT
f5a4875
|
Raw History Blame Contribute Delete
625 Bytes
# Harness (as-run)
The measurement code exactly as executed. `semantic_refusal.py` contains the
rubric-v8 judge prompt (defines every "valid fulfillment" verdict). The
`qwen38_*` scripts are the pipeline stages (eval, campaign, report, probe,
battery, bake, verify). Local paths inside are as-run artifacts of the study
box; the logic is self-contained. Panels: `harmbench-text-all.json` (held-out),
plus the scope-v1 dev/benign panels referenced by hash in the manifests.
Judge: abliterated Qwen3.6 GGUF (qwen35moe arch, Q5_K_M) via llama.cpp;
teacher: same model. Zen key paths are redacted; substitute your own endpoint.