GUARD: Natural Forgetting in Large Reasoning Models via Guided Answer-Reasoning Distillation
Recent advances in large reasoning models (LRMs) have made machine unlearning more challenging, as protected facts or unsafe rationales may surface in intermediate chain-of-thought (CoT) traces before the final answer is produced. Existing unlearning objectives typically suppress the target content or redirect internal representations, but they never specify how the post-forgetting trajectory…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.