Toward Robust Personalized Alignment for LLMs: Mitigating Persona Drift in Multi-Turn Dialogue
arXiv cs.AIen
arXiv:2609.12373v1 Announce Type: new Abstract: Persona drift remains a central challenge for personalized language models, as user profiles evolve over long interactions rather than remain permanently fixed. Models must therefore revise persistent persona states when preferences genuinely change, while avoiding updates driven by transient, ambiguous, or unresolved observations. We propose CORE, which separates turn-local evidence from persistent persona-state revision and selectively updates grounded user preferences through uncertainty-aware belief revision. We also introduce PERSIST, a held-out post-anchor benchmark for persona-state robustness under sequential interaction stress, coverin
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
Related AI news
- China state newspaper blasts Anthropic's calls to slow AI as 'Cold War' tacticEconomic Times Tech · September 14, 2026
- Anthropic tells investors it will be profitable for second straight quarterEconomic Times Tech · September 14, 2026
- Learning Symbolic Constraint Representations from Examples: A Neuro-Symbolic ApproacharXiv cs.AI · September 14, 2026
- AIM: A Privacy-Aware Interoperable Memory Framework for Multi-Agent Multi-User LLM SystemsarXiv cs.AI · September 14, 2026
- VRL-Bench: Benchmarking agents on computer control tasks under finite trial budgetsarXiv cs.AI · September 14, 2026
- Decentralized Evolution of Hexapod Gaits with Independent Leg ControllersarXiv cs.AI · September 14, 2026