Safe to Stop? Risk-Constrained Stopping for Sequential Clinical Diagnosis Agents
arXiv cs.AIen
arXiv:2609.09678v1 Announce Type: new Abstract: Clinical diagnosis agents must decide not only what test to request next, but also when to diagnose or defer. Existing agent benchmarks largely evaluate accuracy after fixed or unconstrained interaction, leaving autonomous stopping reliability implicit. We present Cros, a risk-constrained stopping layer combining state-wise error ranking, policy design on disjoint development splits, and LTT-style exact tests of selective diagnostic error and minimum autonomous coverage for complete sequential policies. Its finite-sample guarantee requires the candidate family, testing rule, and any randomization to be frozen before calibration labels are acces
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Agenter
- Reglering
Related AI news
- Chinese tech giants are hiring skilled professionals as specialized AI trainers to build high-quality datasets, mirroring efforts by US platforms like Mercor (Viola Zhou/Rest of World)Techmeme · September 10, 2026
- Q&A with AI researcher Jacob Coxon, who quit Anthropic, on the need for industry-wide, international coordination to limit recursive self-improvement, and more (Maxwell Zeff/Wired)Techmeme · September 10, 2026
- Anzeige: Ansible f�r automatisiertes SystemmanagementGolem.de · September 10, 2026
- Samsung SDS partners with OpenAI and Anthropic in AI pushDIGITIMES · September 10, 2026
- DeepSeek's next AI test is not the model; it's everything around itDIGITIMES · September 10, 2026
- Meta share price surges after personal AI agent Muse releaseEconomic Times Tech · September 10, 2026