A Four-Stage Decomposition of Word-Problem Solving and Mechanistic Fragility in LLM Math Reasoning
arXiv cs.AIen
arXiv:2609.17804v1 Announce Type: new Abstract: Large language models solve grade-school math word problems with high accuracy, yet a single irrelevant clause inserted into the problem can collapse it. We reconcile these observations with a mechanistic account. We show that the model's internal computation decomposes into a four-stage sequential pipeline, Schema Abstraction, Operation Planning, Operand Binding, and Computation, each stage producing a distinct intermediate representation in an identifiable band of layers. Using the same scaffold to diagnose distractor-induced failure, we localize the corruption to a single stage, Operation Planning, implemented by a set of attention heads who
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
Related AI news
- Sources: Emulate, a month-old UK AI startup founded by former Google DeepMind researchers, is in advanced talks to raise as much as $700M at a $3.7B valuation (Financial Times)Techmeme · September 17, 2026
- The Other Half of the Memory Wall: Serving 35B MoEs from SSD with Trained Routing PredictionarXiv cs.AI · September 17, 2026
- EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading AgentsarXiv cs.AI · September 17, 2026
- SNOMED CT Concept Recommendation from Masked Clinical ContextarXiv cs.AI · September 17, 2026
- Memory Has Geometry: Non-Uniform Geometric Memory for Long-Horizon Personalized AIarXiv cs.AI · September 17, 2026
- When to Call an LLM: A Confidence-Gated Hybrid for Cost-Effective Emotion Recognition in Conversational AIarXiv cs.AI · September 17, 2026