PulseAugur
EN
LIVE 01:19:28

Google DeepMind releases DiffusionGemma for faster text generation

Google DeepMind has released DiffusionGemma, an experimental open-source model that generates text using a diffusion process rather than sequential token-by-token generation. This approach allows for significantly faster text output, up to four times quicker on GPUs, by processing entire blocks of text simultaneously. While the output quality is lower than traditional autoregressive models like Gemma 4, DiffusionGemma is optimized for speed-critical, interactive local workflows and fits within 18GB of VRAM when quantized. AI

IMPACT Accelerates local inference for interactive AI applications by enabling significantly faster text generation.

RANK_REASON Google DeepMind released a new experimental model, DiffusionGemma, with a novel text generation approach.

Read on Google DeepMind →

AI-generated summary · Google Gemini · from 14 sources. How we write summaries →

Google DeepMind releases DiffusionGemma for faster text generation

COVERAGE [14]

  1. Google DeepMind TIER_1 English(EN) ·

    DiffusionGemma: 4x faster text generation

  2. arXiv cs.LG TIER_1 English(EN) · Ali Asaria, Tony Salomone, Deep Gandhi ·

    Neither Parallel Nor Sequential: How DiffusionGemma Actually Commits Tokens

    arXiv:2606.14620v1 Announce Type: new Abstract: Open diffusion language models are marketed as parallel, non-autoregressive decoders, yet the order in which a shipped checkpoint actually commits its tokens is almost never measured. We instrument DiffusionGemma 26B, a masked discr…

  3. arXiv cs.LG TIER_1 English(EN) · Deep Gandhi ·

    Neither Parallel Nor Sequential: How DiffusionGemma Actually Commits Tokens

    Open diffusion language models are marketed as parallel, non-autoregressive decoders, yet the order in which a shipped checkpoint actually commits its tokens is almost never measured. We instrument DiffusionGemma 26B, a masked discrete-diffusion mixture-of-experts model built on …

  4. arXiv cs.CV TIER_1 English(EN) · Bingxuan Dai, Hongsong Wang, Jie Gui ·

    Property-Informed Diffusion-Based Text-to-Microstructure Generation

    arXiv:2606.08150v1 Announce Type: new Abstract: Designing 3D metamaterial microstructures that meet the intended functions remains a major challenge, as it typically requires domain expertise, iterative simulations, and extensive manual tuning. Existing work on inverse design tha…

  5. The Decoder TIER_1 English(EN) · Jonathan Kemper ·

    Google's new open model DiffusionGemma generates text from noise instead of word by word

    <p><img alt="Lettering &quot;DiffusionGemma&quot; in white and blue against a dark blue background with blurred code and text fragments." class="attachment-full size-full wp-post-image" height="1015" src="https://the-decoder.com/wp-content/uploads/2026/06/diffusiongemma-01-hero.j…

  6. Hacker News — AI stories ≥50 points TIER_1 Deutsch(DE) · meetpateltech ·

    DiffusionGemma: 4x Faster Text Generation

  7. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    Google AI Releases DiffusionGemma, a 26B MoE Open Model Using Text Diffusion for Up to 4x Faster Generation

    <p>DiffusionGemma is Google DeepMind's experimental 26B open model using text diffusion for up to 4x faster generation on GPUs.</p> <p>The post <a href="https://www.marktechpost.com/2026/06/10/google-ai-releases-diffusiongemma-a-26b-moe-open-model-using-text-diffusion-for-up-to-4…

  8. Towards AI TIER_1 English(EN) · Anna Jey ·

    DiffusionGemma Developer Guide: When Parallel Text Generation Beats Token-by-Token LLMs

    <figure><img alt="" src="https://cdn-images-1.medium.com/max/1024/1*OEZoLNe6ewNnH8-nD23lJg.jpeg" /><figcaption>DiffusionGemma Developer Guide</figcaption></figure><p>Google’s DiffusionGemma is not just another open model to add to your benchmark spreadsheet. It is a sign that tex…

  9. The Register — AI TIER_1 English(EN) ·

    Google's new open-weights model brings image-generation tricks to AI text generation

    Language model builds on diffusion tech to boost output performance by up to 4x, claims Chocolate Factory

  10. Mastodon — sigmoid.social TIER_1 English(EN) · [email protected] ·

    Google DeepMind has released DiffusionGemma, a 26B open model that uses text diffusion instead of autoregressive decoding. The approach generates entire blocks

    Google DeepMind has released DiffusionGemma, a 26B open model that uses text diffusion instead of autoregressive decoding. The approach generates entire blocks of text in parallel, delivering up to 4x faster generation on GPUs than standard models. It fits in 18GB of VRAM and run…

  11. r/LocalLLaMA TIER_1 English(EN) · /u/z_latent ·

    [Talk] Text Diffusion — Google DeepMind's Brendan O’Donoghue

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1u3f9hs/talk_text_diffusion_google_deepminds_brendan/"> <img alt="[Talk] Text Diffusion — Google DeepMind's Brendan O’Donoghue" src="https://external-preview.redd.it/WQ26AnaTZrMVOGqTjXigInCvvP2RzzTVw9MP3AJPX7w…

  12. r/LocalLLaMA TIER_1 English(EN) · /u/beasthunterr69 ·

    DeepMind Just Dropped "DiffusionGemma" — Text Generation via Image-Style Diffusion Model

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1u29mlk/deepmind_just_dropped_diffusiongemma_text/"> <img alt="DeepMind Just Dropped &quot;DiffusionGemma&quot; — Text Generation via Image-Style Diffusion Model" src="https://external-preview.redd.it/pr58Y7_8…

  13. r/LocalLLaMA TIER_1 English(EN) · /u/tevlon ·

    DiffusionGemma: 4x faster text generation

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1u26s8n/diffusiongemma_4x_faster_text_generation/"> <img alt="DiffusionGemma: 4x faster text generation" src="https://external-preview.redd.it/pr58Y7_82aglIdD6VxjGg1ns65Db_tHjts-39jVNm9U.png?width=640&amp;crop…

  14. Mastodon — mastodon.social TIER_1 Polski(PL) · aisight ·

    Google DeepMind presents DiffusionGemma – a model that sculpts entire blocks of text from digital noise instead of predicting words sequentially. In collaboration with

    Google DeepMind prezentuje DiffusionGemma – model, który zamiast przewidywać słowa sekwencyjnie, rzeźbi całe bloki tekstu z cyfrowego szumu. Dzięki współpracy z Nvidią technologia ta osiąga prędkość do 1000 tokenów na sekundę. # si # ai # sztucznainteligencja # wiadomości # infor…