DENSE: Distilling Agent Trajectories into Evidence-Grounded Shortcut Trees for Self-Refinement
arXiv cs.AIen
arXiv:2609.21423v1 Announce Type: new Abstract: Online agent deployments produce abundant execution traces, while task-specific verification and expert annotation are costly to scale. We study how to distill these traces into reusable feedback without post-hoc outcome labels, drawing on their evidence of local progress, recovery, and unfinished requirements. We introduce DENSE (Distilling Evidence from Nested Subtask Executions), which organizes this evidence into evidence-grounded nested shortcut trees. DENSE compresses redundant attempts, reconciles issues across levels using recovery evidence, and summarizes completed branches while expanding unresolved ones, linking reusable progress to
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Agenter
Related AI news
- A researcher used GPT-6 Astra to decipher a WWI German radio transmission from 1918, one of the 50 famous unsolved ciphers listed on a German science blog (prinz)Techmeme · September 21, 2026
- Styr AI-agenter som om de vore anställda – men låtsas inte att de är människorComputer Sweden · September 21, 2026
- GVPO++: Group Variance Policy Optimization for LLM Post-Training and On-Policy DistillationarXiv cs.AI · September 21, 2026
- Driving on Registers, Reasoning on Risk: Risk-Aware Occupancy for Register-Based End-to-End Autonomous DrivingarXiv cs.AI · September 21, 2026
- RBS-Attention: Radius-Bounded Sparse Prefill for Long-Context Large Language ModelsarXiv cs.AI · September 21, 2026
- Decoupling Internal Representational Changes and Causal Importance in Fine-Tuned Large Language ModelsarXiv cs.AI · September 21, 2026