Google Releases DiffusionGemma, a 26B Model That Generates Text Four Times Faster
Google's experimental DiffusionGemma drafts 256-token blocks in parallel, hitting 1,000+ tokens per second on an H100 — at the cost of output quality versus standard Gemma 4.