CRISP: Persistent Concept Unlearning via Sparse Autoencoders Paper • 2508.13650 • Published Aug 19, 2025 • 16
The Geometry of Harmfulness in LLMs through Subconcept Probing Paper • 2507.21141 • Published Jul 23, 2025