Fine-Tuning Fixes Mode Collapse and Over-Dispersion in LLMs
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.16454v1 Announce Type: new Abstract: Recent work by Doshi and Hauser (2024), Bisbee et al. (2024), and Xie et al. (2026) raises concerns that outputs from large language models (LLMs) tend to be under-diverse: they repeat or resemble one another more often than responses from the population they are meant to represent, a phenomenon known as mode collapse. In this work, we show that whether mode-collapse, or its opposite, occurs depends on the specific model and dataset used. Further, with sufficient supervised fine-tuning (SFT) data, LLM output diversity converges toward that of the target distribution from which fine-tuning data are sampled. To quantify this comparison, we measur
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Företag
Related AI news
- Open-weight model developer Arcee AI reaches $1B-plus valuation with new fundingSiliconANGLE · September 17, 2026
- Open-weight model developer Arcee AI reaches $1B-plus valuation with undisclosed Series B fundingSiliconANGLE · September 17, 2026
- ByteDance's Anew Labs reportedly raises US$290M as AI drug discovery gains momentumDIGITIMES · September 17, 2026
- Bain Capital Ventures raises $1.6b for AI fundTech in Asia · September 17, 2026
- Chip equipment and materials suppliers lead India investment pledges ahead of SEMICON India 2026DIGITIMES · September 17, 2026
- Chinese AI developer Z.ai raises year-end ARR outlook to $3bTech in Asia · September 17, 2026