Damage-Aware Bandit Pruning for Vision and Language Transformers
arXiv cs.AIen
arXiv:2609.05448v1 Announce Type: new Abstract: Structured post-training pruning of transformers requires selecting complete functional units whose suppression causes limited degradation. We formulate structured-unit selection for language and vision transformers as a damage-aware multi-armed bandit problem under a fixed candidate-evaluation budget. Attention heads and MLP channel groups are temporarily masked on calibration batches. Paired damage is the masked loss minus the base loss on the same batch, reducing batch-to-batch variation. A smooth bounded reward drives either a UCB-style policy or fractional-Beta Thompson Sampling, and the final mask is constructed sequentially by adding one
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Reglering
- Företag
Related AI news
- Sources: China Securities Regulatory Commission is informally tightening IPO approvals for humanoid startups after a volatile debut by industry leader Unitree (The Information)Techmeme · September 9, 2026
- When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM AgentsarXiv cs.AI · September 9, 2026
- Compiling VGDL into Causal ModelsarXiv cs.AI · September 9, 2026
- RAPID: Reliability-Aware Pair Importance DistillationarXiv cs.AI · September 9, 2026
- PGP-Clinical-TimeKAN: Prior-Guided Joint Probabilistic Forecasting of Clinical TrajectoriesarXiv cs.AI · September 9, 2026
- SCAFFOLD: Self-Improving Web Agents via Recursive Parametric Skill AbstractionarXiv cs.AI · September 9, 2026