The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.16247v1 Announce Type: new Abstract: Large language models sometimes behave in ways resembling human emotional responses, and recent work has identified internal representations that may explain this. We ask whether LLMs represent pain distinctly from fear, sadness, and generic negative valence, and whether this representation functions as pain would be expected to. We build a dataset describing painful situations across five categories: physical, psychological, social, moral, and cognitive. These are paired with controls for fear, negative emotion, negative world states, sadness, non-painful bodily sensation, arousal, numbness, and neutral content. Using denoised difference-in-me
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
Related AI news
- Chip equipment and materials suppliers lead India investment pledges ahead of SEMICON India 2026DIGITIMES · September 17, 2026
- Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?TechCrunch AI · September 16, 2026
- TypeSafe AI exits stealth with $40M to build AI for use by softwareSiliconANGLE · September 16, 2026
- AI environmental concerns build as lawmakers grapple with tech panicAxios · September 16, 2026
- Google Deepmind launches interdisciplinary institute to tackle the big questions around AGIThe Decoder · September 16, 2026
- Better controls clear a path for AI in financeSiliconANGLE · September 16, 2026