ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction
arXiv cs.AIen
arXiv:2608.13622v1 Announce Type: new Abstract: Open-ended real-world interaction admits multiple valid behaviors: an agent may answer directly, ask for clarification, provide progress updates, or confirm before acting. This flexibility breaks a core assumption behind group-based RL: rollouts compared within a group are no longer guaranteed to be behaviorally comparable. As a result, reward-model preferences over interaction style can distort relative advantages and steer optimization toward reward-preferred behaviors rather than context-appropriate ones. We formalize this as a \textit{reward fairness problem} and propose \textbf{ARC} (Advantage Regularization via Conditioning), a training r
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Agenter
Related AI news
- Tekoälypomo antoi potkut ihmistyöntekijälle – 17 myöhästymisen jälkeenTivi · August 17, 2026
- Qwen 3.8 27B shows a 17GB open-weight general purpose model can have long context, effective tool calling, strong vision ability, and competent code generation (Simon Willison/Simon Willison's Weblog)Techmeme · August 17, 2026
- Modular Cognitive Architecture Emerges in Large Language ModelsarXiv cs.AI · August 17, 2026
- No Universal Signal Predicts Sample-Level LLM Regression under Version UpdatesarXiv cs.AI · August 17, 2026
- Cross-Disciplinary Taxonomy and Modeling of Misunderstanding Generation, Amplification, and Detection, from Pragmatics to AI AgentsarXiv cs.AI · August 17, 2026
- Reward Machines for Signal Temporal LogicarXiv cs.AI · August 17, 2026