Keep It CALM: Analyzing the Limits of Global Unsafety in Text-to-Image Generation
arXiv cs.AIen
arXiv:2610.02300v1 Announce Type: new Abstract: Training-free safeguards for text-to-image generation often rely on a reusable safety signal, such as an unsafe direction or global toxic subspace, applied broadly across prompts. We provide a controlled geometric analysis of this global-unsafety assumption and reveal a consistent coverage-selectivity trade-off: compact unsafe subspaces fail to cover heterogeneous unsafe semantics, whereas broader aggregation increasingly distorts safety-adjacent benign prompts. Motivated by this finding, we propose CALM (Counterfactual Adaptive Local Modulation), a training-free safeguard that replaces uniform global removal with prompt-local counterfactual co
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Bild
Related AI news
- Fujitsu nappasi Istekin miljoonahankinnan – Toimittaa Pirkanmaalle tekoälyjärjestelmänTivi · October 5, 2026
- Global AI servers shift production nearshore, slowing direct Taiwan exports to USDIGITIMES · October 5, 2026
- « Je sais qu’on aura toujours besoin d’humains dans ce domaine » : les professions du lien à l’abri d’un remplacement par l’IALe Monde Pixels · October 5, 2026
- Montag: VW-Partner für autonomes Fahren, Fertiger-Druck auf Notebook-Anbieterheise online – KI · October 5, 2026
- DeepSeek Harness challenges Agent lock-in with Claude Code Mods bridge and open plugin architectureDIGITIMES · October 5, 2026
- World Action Modeling with Progressive Visual PlanningarXiv cs.AI · October 5, 2026