Google's DiffusionGemma proves you don't need to train from scratch to build a text diffusion model
The Decoderen

Instead of training a new model from scratch, Google DeepMind retrofitted Gemma 4 into a diffusion model using less than 10 percent of the original training budget. DiffusionGemma generates 256 tokens in parallel instead of one at a time, hitting about 1,500 tokens per second. Quality still trails the original autoregressive model in benchmarks, especially on reasoning tasks. The article Google's DiffusionGemma proves you don't need to train from scratch to build a text diffusion model appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- Verktyg
- Bild
Related AI news
- AI is flooding Britain's employment courts with lawsuitsThe Decoder · August 9, 2026
- OpenAI pauses Astra AI model over critical cybersecurity concernsEconomic Times Tech · August 9, 2026
- AI's energy appetite drives Nvidia and Amazon to pour billions into massive power infrastructureThe Decoder · August 9, 2026
- Google dismantles Deepmind and bets on a fresh start as Hassabis heads for the exitThe Decoder · August 9, 2026
- Google's AI shakeup suggests it may be prioritizing AI diffusion over frontier-model leadership, betting on AI compute as a bigger economic opportunity (Tim O'Reilly/Asimov's Addendum)Techmeme · August 9, 2026
- This Week In AI: Google loses its brightest minds, and Gemini is still a good boy (in bad way)China AI Markets (Google News) · August 8, 2026