Harness Compilation: Which Decisions Should a Small Vision-Language Model Keep?
arXiv cs.AIen
arXiv:2610.11231v1 Announce Type: new Abstract: Small vision-language models may be able to read external evidence yet struggle to obtain it. We introduce Harness Compilation (HC), an offline procedure that adapts the division of work between a frozen small VLM and its external harness. A large teacher uses student execution traces to revise reusable content and control, while a separate validation set selects the deployed harness. Deployment requires neither weight updates nor teacher calls. Across seven visual question-answering settings with students of at most 9B parameters, HC improves scores over bare students by 9.9-23.9 points, averaged over three independent builds per setting. Inte
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
Related AI news
- On the Clock: Towards Punctual and Productive Time-Budgeted AI AgentsarXiv cs.AI · October 9, 2026
- How Narrative Wrapping Affects LLM Refusal: A Cross-Language Benchmark and DefensearXiv cs.AI · October 9, 2026
- Curating Always-Loaded Context for LLM Agents: A Capacitated Assortment Model with Censored FeedbackarXiv cs.AI · October 9, 2026
- AgentHorizon: Evaluating Agentic Judges for Long-Horizon Computer-Use TasksarXiv cs.AI · October 9, 2026
- When Lower Reconstruction Loss Hurts: Distributionally Robust Refinement for Low-Bit LLM QuantizationarXiv cs.AI · October 9, 2026
- Plan-and-Patch: Diffusion Language Models for Agentic PlanningarXiv cs.AI · October 9, 2026