CogGym: Towards Large-Scale Comparative Evaluation of Human and Machine Cognition
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.21259v1 Announce Type: new Abstract: Understanding and modeling human intelligence are parallel goals shared by artificial intelligence (AI) and cognitive science. As AI systems grow increasingly capable, in what ways do model responses resemble human responses, and where do they systematically diverge? The sheer breadth and diversity of the tasks humans can perform and think about pose a challenge for scalable and rigorous comparison between humans and models. We introduce CogGym, a scalable, unified framework grounded in cognitive science for systematically comparing model and human behavior on matched experimental trials. CogGym uses a semi-automated, human-in-the-loop pipeline
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Företag
Related AI news
- A researcher used GPT-6 Astra to decipher a WWI German radio transmission from 1918, one of the 50 famous unsolved ciphers listed on a German science blog (prinz)Techmeme · September 21, 2026
- Hong Kong-based Qupital, which offers cross-border ecommerce financing to SMEs, raised a $300M Series C led by M Capital as it weighs a possible IPO (FinTech Global)Techmeme · September 21, 2026
- China slows humanoid robot IPO rush as hype outruns realityEconomic Times Tech · September 21, 2026
- Styr AI-agenter som om de vore anställda – men låtsas inte att de är människorComputer Sweden · September 21, 2026
- DENSE: Distilling Agent Trajectories into Evidence-Grounded Shortcut Trees for Self-RefinementarXiv cs.AI · September 21, 2026
- GVPO++: Group Variance Policy Optimization for LLM Post-Training and On-Policy DistillationarXiv cs.AI · September 21, 2026