When Do LLMs Apply the Wrong Law? Diagnosing LLM Failures in Temporal Legal Reasoning
arXiv cs.AIen
arXiv:2608.14610v1 Announce Type: new Abstract: Legal reasoning tasks such as legal judgment prediction (LJP) require identifying the temporally correct version of the law governing a case -- a capability we term temporal applicable-law determination. However, whether large language models (LLMs) can reliably perform this task remains unexplored. In this paper, we construct a benchmark to evaluate LLMs on temporal applicable-law determination, and systematically investigate why they fail at temporal legal reasoning. Our experiments reveal four key findings. First, LLMs exhibit a strong bias toward applying the most recently enacted law, regardless of when the legally relevant facts occurred.
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Reglering
Related AI news
- Dienstag: Trump profitiert trotz Sanktionen, Cyberangriff auf Berliner Netzheise online – KI · August 18, 2026
- Global AI Regulations for FAIR and Ethics in High-Risk Use Cases: A Comparative ReviewarXiv cs.AI · August 18, 2026
- From Doyle to AGM: A Survey and an Implementation Roadmap for Belief ChangearXiv cs.AI · August 18, 2026
- Position: Want Better ML Reviews? Stop Asking Nicely and Start Incentivizing with a Credit SystemarXiv cs.AI · August 18, 2026
- Longitudinal and Graph-Augmented Prediction of Adolescent Substance Use Onset in the ABCD StudyarXiv cs.AI · August 18, 2026
- OGX: An Open-Source, Vendor-Neutral Generative AI Application ServerarXiv cs.AI · August 18, 2026