Principled Thoughts for Latent Recursive LLM Systems
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.36159v1 Announce Type: new Abstract: Large language models can reason in continuous space instead of decoded text, by recurring on their own hidden states or by passing those states between agents, while training supervises only the Cross-Entropy (CE) of the final decoded answer and does not constrain the thought. Theoretical and empirical analyses establish and confirm four failures of CE-only training that lead to a lower probability of the correct answer such as collapsing thoughts across distinct questions and retaining irrelevant information. We introduce REST (REpresentation-Supervised Thoughts), a training objective that turns four properties of a valid thought representati
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Agenter
Related AI news
- Hands-on with Dots, OpenAI's work-focused agentic product: highly capable and intuitive, with natural-feeling conversations represented as a chat inside ChatGPT (Casey Newton/Platformer)Techmeme · September 30, 2026
- Chinese firms trail global peers on profits, but AI power boom offers bright spot: NatixisSCMP Tech · September 30, 2026
- More Features Are Not More Evidence: Limits of Training-Free Human Activity Recognition with JevarXiv cs.AI · September 30, 2026
- Snart här: AI-datorer som sänker tokenkostnadernaComputer Sweden · September 30, 2026
- Towards Mitigating Deceptive Safety Alignment in Large Reasoning ModelsarXiv cs.AI · September 30, 2026
- GeoOutageBench: Benchmarking Ambiguity-aware, Ontology-grounded Geospatiotemporal KGQA for Multimodal Power Outage and Resilience AnalysisarXiv cs.AI · September 30, 2026