← Back to Model Beat
Models·Aug 9·all news from August 9, 2026

Google's DiffusionGemma proves you don't need to train from scratch to build a text diffusion model

Google DeepMind has adapted the Gemma 4 language model into a diffusion model, allowing it to generate text in parallel rather than token by token. By using less than 10 percent of the original training budget, this technique achieves speeds of approximately 1,500 tokens per second. This development demonstrates that existing large language models can be repurposed for faster text generation without requiring entirely new training processes.

Covered by 1 source

Related stories

ModelsDeepSeek Jacks Up Price for Flagship AI Models Ahead of IPOAug 10 · 34 sourcesModelsAlibaba AI Models Hit 3 Billion Downloads, Passing Meta, GoogleAug 13 · 32 sourcesModelsIntroducing Gemini 3.7 FlashAug 11 · 13 sourcesModelsMeta is back with Muse Glimmer: local, agentic, multimodal, and open sourceAug 10 · 31 sources