The Independence Prior of SAEs Fragments Visual Concepts
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2610.04112v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) decompose model activations into sparse combinations of interpretable dictionary atoms. Although SAEs are grounded in the Linear Representation Hypothesis (LRH), their objective smuggles in an additional prior: concepts across patches are treated as independent, an assumption clearly violated by natural images and by the activations they induce. We therefore specialize LRH to vision through the Markov-Field Linear Representation Hypothesis (MFLRH), which adds the missing spatial dependencies to the LRH assumptions. We thus propose Spatial-SAE as an amortized MAP estimator under the MFLRH. Spatial-SAE consistently outp
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Bild
Related AI news
- Huawei's Kirin 9050 Pro reveals new logic folding chip designDIGITIMES · October 6, 2026
- Training Numerical Intelligence via Auto-Diagnosis and Skill DiscoveryarXiv cs.AI · October 6, 2026
- CUAWright: A Minimal Unified Interface for Digital AgentsarXiv cs.AI · October 6, 2026
- InvestigationWorlds: An Agentic Environment for Legal InvestigationarXiv cs.AI · October 6, 2026
- Auditing Pairwise Equivalence Judgments: Self-Critique Effects and Diversity Measurement in Multi-Agent Hypothesis GenerationarXiv cs.AI · October 6, 2026
- Agentic Cognitive Depth: Operational Criteria for Evaluating LLM AgentsarXiv cs.AI · October 6, 2026