safetycritiquebearishExisting unlearning benchmarks evaluate solely at the output level, leaving open whether unlearning truly erases knowledge or merely obfuscates itMachine Learning27 Jul 2026http://arxiv.org/abs/2607.02513v1