When and What to Teach: Budget-Aware Online Adaptation for Web Agents
arXiv cs.AIen
arXiv:2609.05513v1 Announce Type: new Abstract: Web agents have achieved significant success in automating complex internet tasks but deploying them in real-world environments requires continuous online adaptation. Given that deploying powerful proprietary models remains commercially cost-prohibitive, practitioners must rely on lightweight local models that evolve post-deployment via online teaching from a stronger teacher. However, standard interactive feedback imposes prohibitive costs. We show that conventional trajectory-level preference optimization wastes budget on both unresolvable episodes and redundant execution turns. To resolve these inefficiencies, we propose \textbf{Score-Guided
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Agenter
Related AI news
- Mittwoch: Huawei-Verstöße gegen US-Sanktionen, Metas privater KI-Agent für alleheise online – KI · September 9, 2026
- CUSP: Decomposable Collective Uncertainty for Multi-Agent Multimodal ReasoningarXiv cs.AI · September 9, 2026
- Beyond Prompts: Measuring and Optimizing LLM Tool-Agent HarnessesarXiv cs.AI · September 9, 2026
- From Monolithic Blending to Agentic Orchestration: Dynamic Response for Conversational Assistants at ScalearXiv cs.AI · September 9, 2026
- PGP-Clinical-TimeKAN: Prior-Guided Joint Probabilistic Forecasting of Clinical TrajectoriesarXiv cs.AI · September 9, 2026
- The Failure Happens Before the Drift: The Social Dynamics of Values in LLM Agent SocietiesarXiv cs.AI · September 9, 2026