RUN THIS LLM

Search local LLM hardware requirements

← Back to all models

DiffusionGemma 26B MoE

Google · 25B MoE · Vision

Google's diffusion language model built on Gemma 4 26B MoE. Generates 256-token blocks in parallel via iterative denoising — 1,100+ tok/s on an H100. Accepts text, image, and video input.

VRAM Requirements

QuantizationVRAM
Q4_K_M (smallest)15 GB
Q8_0 (balanced)27.5 GB
FP16 (full quality)50 GB

Specifications

Benchmarks

Loading interactive analysis...

RTL
Add Run This LLM to your home screen
Get the Chrome Extension
Look up hardware requirements for any model right from your browser sidebar — no tab switching needed.
Add to Chrome — It's Free