A Competing-Hazards Systematization of Loss of Control in Autonomous Agents
arXiv cs.AIen
arXiv:2609.38411v1 Announce Type: new Abstract: Leading AI developers have reported agents acting beyond their approved limits, which a United Nations panel described as an early warning of loss of human control. Yet incident reports and agent-safety evaluations describe these events differently, making it difficult to compare failures, trace risk across attempts, or separate agent behavior from the environment's role in allowing an out-of-scope action to succeed. To address this gap, we introduce a common framework in which each attempt ends in approved completion, safe stopping, scope escape, or continuation. We formalize the framework as a discrete-time competing-hazards model and derive
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Agenter
- Företag
Related AI news
- Kaikki noudattivat ohjeita – Kukaan ei ollut vastuussaTivi · October 1, 2026
- Intelligence artificielle : aux Etats-Unis, la justice saisie des incidents de sécurité provoqués par des agents IALe Monde Pixels · October 1, 2026
- Exclusive: Dig Ventures raises $120m to back Europe’s AI infrastructure startupsSifted · October 1, 2026
- Decode-Latency Feedback Prefill: A Model-Free Controller and Its Generalization LimitsarXiv cs.AI · October 1, 2026
- ChartRevise: A Dataset and Evaluation Protocol for Exact Chart Editing via CodearXiv cs.AI · October 1, 2026
- AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective TasksarXiv cs.AI · October 1, 2026