Google AI Releases DiffusionGemma, a 26B MoE Open Model Using Text Diffusion for Up to 4x Faster Generation
4x faster. Google's new DiffusionGemma uses text diffusion to ship a 26B MoE model that rewrites the inference speed playbook.

Why it matters
Google DeepMind is challenging the autoregressive inference bottleneck with a diffusion-based approach to text generation, offering significant speed gains that could reshape how teams think about model deployment and inference cost.
The key facts
6 to know26B parameters with Mixture-of-Experts architecture
Up to 4x faster generation on GPUs
Text diffusion methodology (non-autoregressive)
Open model release
Google DeepMind source
Experimental status
Go to the source
MarkTechPostmarktechpost.com
Publisher excerpt: DiffusionGemma is Google DeepMind's experimental 26B open model using text diffusion for up to 4x faster generation on GPUs.