Google DeepMind retrofits Gemma 4 into a parallel-token diffusion model, DiffusionGemma
Google DeepMind converted Gemma 4 into a text-diffusion model using under 10% of the original training budget, generating 256 tokens at once at ~1,500 tokens/sec.