Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2610.00531v1 Announce Type: new Abstract: AI-generated content, often called AI slop, is increasingly common everywhere, particularly in academia. Slop in AI-generated scientific papers, however, has more complex patterns that cannot be easily detected by existing token-based AI detectors. Each part of such a paper looks plausible while the scientific reasoning that connects the parts breaks down, which can mislead how readers assess the work. We benchmark these failures as scientific slop through six measures across Structure, Argument, and Artifacts. We construct SciSlopBench with 390 AI-generated papers, mostly in computer science but spanning the life, social, and natural sciences,
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
Related AI news
- ArXiv limits preprint submissions to two per month per submitter, as AI access fuels a record 40,363 submissions in September 2026, vs. 20,569 in September 2024 (Kat Boboris/arXiv)Techmeme · October 2, 2026
- Google's first Suncatcher satellite reaches orbit to test AI chips in spaceDIGITIMES · October 2, 2026
- Before Agents Decide: Epistemic Action in LLM-Based SystemsarXiv cs.AI · October 2, 2026
- Heavy-Tailed Memory Traces in Long-Horizon Language AgentsarXiv cs.AI · October 2, 2026
- Measuring the Microtask Eligibility Gap: When Is an Off-the-Shelf SLM Enough for an Agent Harness?arXiv cs.AI · October 2, 2026
- EviGraph: Proof-Carrying Selective Recommendation over Temporal Public-Service Knowledge GraphsarXiv cs.AI · October 2, 2026