Spectral Surgery β Commonsense Reasoning
Collection
Qwen3-8B adapters trained on Commonsense170K, with post-hoc Spectral Surgery evaluation across eight commonsense reasoning benchmarks. β’ 9 items β’ Updated
YAML Metadata Warning:empty or missing yaml metadata in repo card
Check out the documentation for more information.
Post-hoc Spectral Surgery applied to a Qwen3-8B LoRA adapter fine-tuned on Commonsense170K.
No additional gradient-based training is performed during Spectral Surgery.
o_proj, down_projSee spectral_edit_meta.json for the exact edit metadata.
| Task | LoRA | + Spectral Surgery |
|---|---|---|
| BoolQ | 88.0734 | 88.2875 |
| PIQA | 90.2067 | 90.4788 |
| SocialIQA | 82.2416 | 82.2927 |
| HellaSwag | 94.2243 | 94.1844 |
| WinoGrande | 89.4238 | 89.5028 |
| ARC-Easy | 97.1801 | 97.3906 |
| ARC-Challenge | 91.5529 | 90.7850 |
| OpenBookQA | 93.4000 | 93.4000 |
| Macro | 90.7879 | 90.7902 |
| Micro | 91.8373 | 91.8640 |
Correct predictions:
On this eight-task commonsense suite, Spectral Surgery preserves the aggregate performance of the source LoRA adapter.
The small numerical difference should not be interpreted as a meaningful improvement.