Skip to content
https://abc.microfintool.com/

Mobile & Accessories

  • Welcome to ABC Tool: Your Ultimate Portal for Smartphones and Accessories
  • About / Contect
    • PRIVACY POLICY
  • Blog
Google’s latest DiffusionGemma open AI model comes with a 4x speed boost

Google’s latest DiffusionGemma open AI model comes with a 4x speed boost

Posted on June 10, 2026 By safdargal12 No Comments on Google’s latest DiffusionGemma open AI model comes with a 4x speed boost
Blog

[ad_1]

Another day, another AI model from Google. This time, Google DeepMind has released a new member of the Gemma 4 open model family, but it’s fundamentally different from the rest of the lineup. DiffusionGemma doesn’t generate outputs linearly like most AI models. Instead, it can produce an entire block of text in parallel. Google says this makes it faster and more efficient when running on local hardware like an Nvidia DGX or a humble gaming GPU.

Most AI models are designed to be autoregressive—they generate text left to right one token at a time. DiffusionGemma has more in common with image generation models, which start with static and then denoise it to create the desired content. This model takes a field of placeholder tokens running over the canvas multiple times to generate likely tokens and using those to improve estimation of others. At the end of the process, the model finalizes its token outputs in one large block—the “denoised” text canvas.

DiffusionGemma is fairly large in the realm of Google’s open models. It’s a Mixture of Experts (MoE) model with a total of 26 billion parameters, but only 3.8 billion are activated during inference. That means it should fit in the 18GB ram allotment of a high-end GPU. In testing with an RTX 5090, DiffusionGemma spits out around 700 tokens per second. With a single Nvidia H100 AI accelerator, DiffusionGemma can produce 1,000+ tokens per second. That’s about four times the output of the similarly sized autoregressive Gemma models.

This approach to text generation shifts the bottleneck from memory bandwidth to compute, generating up to 256 tokens in parallel. Google says this offers a measurable boost in non-linear tasks like in-line editing, molecular sequencing, and mathematical graphing. The animation above shows how DiffusionGemma was tuned to solve Sudoku puzzles, which is a notoriously challenging task for standard autoregressive AI models because each token depends on future tokens. DiffusionGemma’s ability to continuously self-correct large sets of tokens makes that easier.

[ad_2]

Source link

Post Views: 61

Post navigation

❮ Previous Post: North Koreans behind nearly half of US tech industry hacks, says CrowdStrike
Next Post: My Eyes Love Logitech’s New Mobi Fold Mouse, My Hand a Little Less So ❯

You may also like

Good news for perfectionists with a Kindle Scribe
Blog
Good news for perfectionists with a Kindle Scribe
April 18, 2026
Life Got So Much Better When I Turned Off My Phone Notifications
Blog
Life Got So Much Better When I Turned Off My Phone Notifications
June 8, 2026
Samsung's memory division posts massive profits for Q1, smartphones still profitable
Blog
Samsung's memory division posts massive profits for Q1, smartphones still profitable
May 1, 2026
American Airlines Signs Up for Starlink Wi-Fi Service on Its Flights
Blog
American Airlines Signs Up for Starlink Wi-Fi Service on Its Flights
May 27, 2026

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Recent Posts

  • Fix Outlook for Mac 16.110 Missing Email History
  • Anthropic’s Claude Tag: Smarter Slack Assistant
  • Google Home Facial Recognition Update Boosts Accuracy
  • Meta Pauses Employee Tracking After Data Leak
  • US Accelerates Post-Quantum Cryptography Deadline to 2030

Recent Comments

  1. nszgpwrtqv on WhatsApp is now testing its subscription service, here's what you get and how much it costs
  2. qmkffkjqvx on Man dies covered in necrotic lesions after amoebas eat him alive
  3. Declan Chidlow on History of Game Console Web Browsers: Evolution & Tech
  4. ALEXAnync on Meta steals a tactic from Tesla and builds data centers in tents
  5. ALEXAnync on The Delivery You Didn’t Order: Breaking Down the ‘Free Phone’ Scam

Archives

  • June 2026
  • May 2026
  • April 2026

Categories

  • Blog

Copyright © 2026 Mobile & Accessories.

Theme: Oceanly News by ScriptsTown