Is DiffusionGemma (Google) Down?

Quickest ways to check the current status of blog.google plus recent known issues and working alternatives if it's out.

Last editorial review: 2026-06-10

How to check right now

Known issues we've tracked

LAUNCH (2026-06-10): DiffusionGemma released as an experimental open model. 26B-total / 3.8B-active MoE, Apache 2.0, text-diffusion architecture generating 256-token blocks in parallel with bidirectional attention. Vendor-claimed speeds: up to 4x faster than comparable autoregressive models, 1,000+ tok/s on H100, 700+ tok/s on RTX 5090, ~18GB VRAM quantized. Availability: Hugging Face weights, NVIDIA NIM, Gemini Enterprise Model Garden; vLLM/MLX/llama.cpp support at launch. Google's own caveats: trails Gemma 4 on quality benchmarks; speed advantage weaker on Apple Silicon and in high-QPS serving. Front-page Hacker News reception (226 points) on launch day
2026-06-10Google blog (blog.google/innovation-and-ai/technology/developers-tools/diffusion-gemma-faster-text-generation/), Google developers blog (developers.googleblog.com/en/diffusiongemma-the-developer-guide/), NVIDIA blog (RTX AI Garage)

Issues here are sourced from our editorial sweeps, not real-time telemetry. Newer issues may exist.

What to use if DiffusionGemma (Google) is down

Top AIToolTier-ranked alternatives in the same category, ordered by our overall score.

A

Qwen (Alibaba)

8.8

Alibaba's open-weights + API family, and August 2026 was its biggest open-release month yet: **Qwen3.8-2.4T-A95B -- the Max-class flagship -- went open-weight on 2026-08-08** (custom Qwen3.8-Max licence), **Qwen3.8-27B** dense landed Apache 2.0 (8/05), and **Qwen3.8-Flash-Next** (8/24 weights, 8/26 blog) previews the Qwen4 architecture: 125B main + 51B n-gram embeddings with only 6B active, served as Qwen3.8-Flash at $0.15/$0.47 per 1M. Qwen 3.7 Max remains the GA API flagship ($2.50/$7.50)

A

MiniMax M3

8.4

**MiniMax-H3's weights are public on Hugging Face (repo 2026-07-28, updated 2026-08-13) under a community licence whose open-weight grant is limited to the US, EU, UK and South Korea**, and **MiniMax Music 3 (2026-08-07)** ships open weights for five-minute full-song generation with a prominent-attribution commercial clause. MiniMax's coding/agent flagship -- M3 (June 1 2026): 1M-token context, MSA sparse attention (>15x decoding speedup at long context), SWE-Bench Pro 59.0%, Terminal-Bench 66.0%. OPEN WEIGHTS LIVE on HuggingFace since June 12 (~428B total / ~23B active, native multimodal, minimax-community license)

A

Gemma 4 (Google)

8.3

Google DeepMind's open-weights model family -- multimodal, 256K context, runs on edge devices

A

IBM Granite 4.0

8.2

IBM's enterprise-focused open-weight family -- Granite 4.0 hybrid Mamba-2 + transformer architecture (70-80% memory reduction vs pure transformer), 3B to 32B sizes, Apache 2.0. First open model family to secure ISO 42001 certification. Nano 350M runs on CPU with 8-16GB RAM. 3B Vision variant landed 2026-04-01

A

Kimi K3 (Moonshot)

8.1

Moonshot's 2.8T-parameter Kimi K3 (launched 2026-07-16/17) is the largest open-weight model ever released -- 1M context, multimodal, $3/$15 per 1M via API, ranked best-available on Arena.AI at launch. WEIGHTS SHIPPED ~2026-07-26/27 on Hugging Face (2.8T total / 104B activated, safetensors) under a custom Kimi K3 License, not the Modified MIT of the K2 line

About DiffusionGemma (Google)

Tier C (6.8/10). Google DeepMind's experimental open-weights TEXT-DIFFUSION model (June 10, 2026) -- 26B MoE (3.8B active), Apache 2.0, generates 256-token blocks in parallel with bidirectional attention for up to 4x faster output (1,000+ tok/s on H100). Trades some quality vs Gemma 4 for raw speed