TPvG: A Moral Decision Framework for Large Language Models from One-Shot to Sequential Feedback
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2608.28610v1 Announce Type: new Abstract: Existing LLM moral evaluations typically present models with isolated moral vignettes and elicit a single-shot decision, neglecting a factor known to profoundly influence human moral behavior: consequence feedback. We introduce TPvG (Text-based Pain-versus-Gain), adapted from a human moral paradigm, which embeds consequence feedback into an everyday moral dilemma of not harming others versus maximising self-gain. TPvG comprises five moral decision tasks, progressing from minimal-context one-shot choices to sequential decisions with explicit consequence feedback. Our results show that LLM moral decisions were strongly affected by decision format
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Företag
Related AI news
- SoftBank, Nvidia, OpenAI-backed SB Energy files for US IPOTech in Asia · September 2, 2026
- Perplexity CEO announces rollout of hybrid compute feature for Mac applicationEconomic Times Tech · September 2, 2026
- Alibaba Cloud backs Malaysia’s AI untuk RakyatTech in Asia · September 2, 2026
- John Ternus takes over as Apple enters era of AI, foldable phonesEconomic Times Tech · September 2, 2026
- Source: OpenAI's Astra model uses "recurrent depth", a technique that improves cost and performance but obscures the AI's reasoning, making it harder to monitor (The Information)Techmeme · September 2, 2026
- Anthropic launches Claude Fable 5.1 and Mythos 5.1, cuts agentic-task costs by up to 45%DIGITIMES · September 2, 2026