REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs
Apple Machine Learningen
Apple Machine Learning
AI Global WireMost current vision-language-action (VLA) models—such as OpenVLA, π0, RT-2, and RDT-1B—are “monolithic.” This means they generate raw motor commands or very short sequences of actions, without organizing behaviors into reusable, well-defined abstractions. As a result, these models perform poorly on long-horizon (multi-step) tasks, and it’s difficult to interpret what they have learned. Existing approaches for discovering skills often avoid the core problem of deciding when two action sequences are “behaviorally equivalent.” For example, AtomicVLA and AtomSkill group action sequences by…
This is a short summary published by AI Global Wire. The full article is owned and hosted by Apple Machine Learning — open it there to read it in full.
Read the full story at Apple Machine Learning- Verktyg
Related AI news
- Broadcom enlists partners to tie VCF adoption to business outcomesSiliconANGLE · September 2, 2026
- We’re ‘dangerously close’ to dead internet theory, says Pangram’s CEOTechCrunch AI · September 2, 2026
- Gemini 3.8 Flash is Google's third budget model in six weeks while frontier models remain MIAThe Decoder · September 2, 2026
- TIME AI 100 spotlights China's new AI guard, from Moonshot, Z.ai to AgiBot, ManusDIGITIMES · September 2, 2026
- Lyte raises $165M at $1.6B valuation to bring accurate perception to robotsSiliconANGLE · September 2, 2026
- Security stops being a layer in the stack and starts governing the AI economySiliconANGLE · September 2, 2026