Leak-Resistant Unlearning: A New Benchmark for Evaluating Multi-Hop Reasoning Consistency and Recovery Robustness
arXiv cs.AIen
arXiv:2608.04519v1 Announce Type: new Abstract: Benchmarking machine unlearning methods is critical to understand whether sensitive knowledge is removed from large language models (LLMs) or not. Current unlearning benchmarks include mainly single-hop questions and a narrow set of multi-hop questions. Although effective, they still face two challenges. (1) Knowledge is not isolated, whereby diverse multi-hop reasoning paths can potentially induce knowledge leakage than normal queries. (2) Unlearning may be fragile: unlearned knowledge can be partially recovered through recovery attacks such as lightweight post-unlearning adaptation, making static evaluation insufficient. Therefore, in this pa
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Företag
Related AI news
- Samsung, SK Hynix shareholders call for bigger payouts from AI cash mountainEconomic Times Tech · August 6, 2026
- Anthropic and OpenAI Agents in soup againEconomic Times Tech · August 6, 2026
- FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM AgentsarXiv cs.AI · August 6, 2026
- Adversarially Robust Abductive Fusion of Pre-trained Transformer-based Perception ModelsarXiv cs.AI · August 6, 2026
- SafeCommit: Certifying When Memory-Grounded Agents May Safely ActarXiv cs.AI · August 6, 2026
- Improving Auto-Design of Neural PDE Solvers with a Domain-Specific LanguagearXiv cs.AI · August 6, 2026