+
+

Related Products

  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Google AI Studio
    30 Ratings
    Visit Website
  • Checksum.ai
    1 Rating
    Visit Website
  • Adobe Firefly
    25,029 Ratings
    Visit Website
  • Google Cloud Speech-to-Text
    366 Ratings
    Visit Website
  • Runpod
    220 Ratings
    Visit Website
  • Imorgon
    5 Ratings
    Visit Website
  • MEXC
    188,765 Ratings
    Visit Website
  • EBizCharge
    207 Ratings
    Visit Website
  • CallTrackingMetrics
    937 Ratings
    Visit Website

About

DiffusionGemma is an experimental open model that explores text diffusion, an exceptionally fast approach to text generation. Released under an Apache 2.0 license, this 26B Mixture of Experts (MoE) model moves beyond the sequential token-by-token processing of typical autoregressive Large Language Models (LLMs). Instead, it generates entire blocks of text simultaneously, delivering up to 4x faster text generation on GPUs. Built on the intelligence-per-parameter of the Gemma 4 family and Gemini Diffusion research, DiffusionGemma integrates a novel diffusion head designed to maximize generation speed. It is designed for researchers and developers exploring speed-critical, interactive local workflows such as in-line editing, rapid iteration, and non-linear text structures. By shifting the decode bottleneck from memory bandwidth to compute, it can generate more than 1,000 tokens per second on a single NVIDIA H100 and more than 700 tokens per second on an NVIDIA GeForce RTX 5090.

About

Google AI Edge Gallery is an experimental, open source Android app that demonstrates on-device machine learning and generative AI use cases, letting users download and run models locally (so they work offline once installed). It offers several features including AI Chat (multi-turn conversation), Ask Image (upload or use images to ask questions, identify objects, get descriptions), Audio Scribe (transcribe or translate recorded/uploaded audio), Prompt Lab (for single-turn tasks such as summarization, rewriting, code generation), and performance insights (metrics like latency, decode speed, etc.). Users can switch between different compatible models (including Gemma 3n and models from Hugging Face), bring their own LiteRT models, and explore model cards and source code for transparency. The app aims to protect privacy by doing all processing on the device, no internet connection needed for core operations after models are loaded, reducing latency, and enhancing data security.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

AI researchers building low-latency local applications who need faster experimental text generation for interactive workflows

Audience

Researchers, and power users interested in a tool to try out or build AI/GenAI models locally

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Google
Founded: 1998
United States
blog.google/innovation-and-ai/technology/developers-tools/diffusion-gemma-faster-text-generation/

Company Information

Google
United States
github.com/google-ai-edge/gallery/

Alternatives

Mercury 2

Mercury 2

Inception

Alternatives

Gemma 3n

Gemma 3n

Google DeepMind
Gemini Diffusion

Gemini Diffusion

Google DeepMind
Mercury Coder

Mercury Coder

Inception Labs
LFM2

LFM2

Liquid AI
ByteDance Seed

ByteDance Seed

ByteDance
LiteRT

LiteRT

Google

Categories

Categories

Integrations

Gemini Enterprise Agent Platform
Gemma
Gemma 3n
Hugging Face
LiteRT
NVIDIA NIM

Integrations

Gemini Enterprise Agent Platform
Gemma
Gemma 3n
Hugging Face
LiteRT
NVIDIA NIM
Claim DiffusionGemma and update features and information
Claim DiffusionGemma and update features and information
Claim Google AI Edge Gallery and update features and information
Claim Google AI Edge Gallery and update features and information