azukivc's picture jenerallee78's picture
Duplicate from jenerallee78/Qwen3.8-27B-Abliterated-SFT
f5a4875
|
Raw History Blame Contribute Delete
625 Bytes

Harness (as-run)

The measurement code exactly as executed. semantic_refusal.py contains the rubric-v8 judge prompt (defines every "valid fulfillment" verdict). The qwen38_* scripts are the pipeline stages (eval, campaign, report, probe, battery, bake, verify). Local paths inside are as-run artifacts of the study box; the logic is self-contained. Panels: harmbench-text-all.json (held-out), plus the scope-v1 dev/benign panels referenced by hash in the manifests. Judge: abliterated Qwen3.6 GGUF (qwen35moe arch, Q5_K_M) via llama.cpp; teacher: same model. Zen key paths are redacted; substitute your own endpoint.