ChronoSRL: Temporal Geometry for Self-Supervised Reinforcement Learning
arXiv cs.AIen
arXiv:2609.36238v1 Announce Type: new Abstract: A goal that is close in space can be far away in time. Obstacles, terrain, and the agent's own capabilities determine how long it takes to get there. Yet, critics in contrastive and survival reinforcement learning do not measure the distances in their representation space in units of time. We therefore introduce ChronoSRL, which gives the critic's embeddings an explicit temporal geometry. The distance between state-action and goal embeddings is trained to match the time that the agent takes to reach the goal (goal-reaching time), while goals that were not reached, and goals from other trajectories, are pushed at least one discount horizon away.
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Agenter
Related AI news
- Hands-on with Dots, OpenAI's work-focused agentic product: highly capable and intuitive, with natural-feeling conversations represented as a chat inside ChatGPT (Casey Newton/Platformer)Techmeme · September 30, 2026
- Chinese firms trail global peers on profits, but AI power boom offers bright spot: NatixisSCMP Tech · September 30, 2026
- Snart här: AI-datorer som sänker tokenkostnadernaComputer Sweden · September 30, 2026
- GeoWind2Plan: Mission-Time 3D Urban Wind Prediction for Energy-Efficient UAV PlanningarXiv cs.AI · September 30, 2026
- An Exact Generate - Transform Decomposition of Small-LLM Team Scaling Across Orchestration ArchitecturesarXiv cs.AI · September 30, 2026
- The Layer Mystery of VLA: An Information-Theoretical Analysis of VLA Latent InterfacearXiv cs.AI · September 30, 2026