Lexsi/audit-recover-apply_safe_lora-llama31-8b-code
8B • Updated • 3
Frontier research around Safe and aligned intelligence
Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution
$C$-$ΔΘ$: Circuit-Restricted Weight Arithmetic for Selective Refusal