Proof-Carrying Cognition: Closing the Verification Gap with Reality-Settled Reward
arXiv cs.AIen
arXiv:2609.09776v1 Announce Type: new Abstract: Frontier gains in language-model reasoning come from reinforcement learning on reasoning traces and are concentrated in domains with a cheap, sound verifier. We argue the field's binding constraint is the verification gap: no scalable, incorruptible reward for reasoning outside formal domains. We make four contributions. (1) Theory: in a joint-Gaussian model of best-of-N selection, verifier-gold correlation rho is the exact exchange rate between test-time compute and capability, and an unsound verifier pays a polynomial penalty N^(1/rho^2); a margin-free copula form predicts realized soundness of real LLM judges to 4% median error. (2) Demonstr
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
Related AI news
- Q&A with AI researcher Jacob Coxon, who quit Anthropic, on the need for industry-wide, international coordination to limit recursive self-improvement, and more (Maxwell Zeff/Wired)Techmeme · September 10, 2026
- Generative AI a new tool in Mali's information war: studyEconomic Times Tech · September 10, 2026
- RobustSGPO: Search-Space Control for Agent Harness EvolutionarXiv cs.AI · September 10, 2026
- Which Tokens Should SFT Actually Learn? A Token-Trimming Perspective on Mathematical ReasoningarXiv cs.AI · September 10, 2026
- A Function-Space Approach to the Statistical Mechanics of Learning DynamicsarXiv cs.AI · September 10, 2026
- Black-Box Red Teaming of Agentic AI: A Taxonomy-Driven Framework for Automated Risk DiscoveryarXiv cs.AI · September 10, 2026