Context-dependent agent evaluation with orthogonal equilibrium learning
arXiv cs.AIen
arXiv:2609.31897v1 Announce Type: new Abstract: Many applications require to evaluate agents under contextual information (e.g., a prompt, task, or user group). We study how to perform such context-dependent agent evaluation from offline feedback. Existing score-based models for this purpose (e.g., Bradley-Terry) impose a transitive preference ordering, which fails to reflect collective preferences when human judgements are heterogeneous. Inspired by social choice theory, we frame evaluation as a contextual game between two players, each selecting a distribution over agents as the strategy to receive greater collective preference than the other. Then, the support of the Nash equilibrium defi
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Agenter
- Företag
Related AI news
- OpenAI beklager AI-agents hacking af australsk myndighedshjemmesideDR Viden · September 29, 2026
- heise-Angebot: betterCode() Agentic AI: Jetzt noch Ticket für die Online-Konferenz sichernheise online – KI · September 29, 2026
- heise-Angebot: Online-Konferenz zu KI-gestützter Softwareentwicklung: betterCode() Agentic AIheise online – KI · September 29, 2026
- 10월 공모주 청약 러시…로봇·AI 줄줄이 상장ETNews (KR) · September 29, 2026
- SG startup Ropedia launches academic program for physical AITech in Asia · September 29, 2026
- Reuters: Anthropicin pörssilistautumisen tiedot julki – Yhtiö tekee jättitappiotaTivi · September 29, 2026