Gemma 4 Models — Hardware Requirements

25 Gemma 4 models from Google and the community, from the smallest that runs in 2 GB of VRAM up to 32.7B parameters. Every row links to full quantization tables and GPU compatibility.

All Gemma 4 Models by Size

ModelParamsContext
Gemma 4 E2B IT Qat Mobile Transformers2.3B131K
Gemma 4 E4B IT Assistant4B131K
OFFELLIA Gemma 4 E4B 8B Claude 4.6 Opus Reasoning MTP4B
Turkish Gemma 4B T1 Scout4.3B131K
Gemma 4 E2B IT Qat Q4 0 Unquantized5.1B131K
Gemma 4 E2B IT Qat Q4 0 Unquantized Heretic5.1B131K
Gemma 4 E2B IT5.1B131K
Gemma 4 E2B IT Uncensored5.1B131K
Supergemma4 E4b Abliterated7.5B131K
Gemma 4 E4B IT Qat Q4 0 Unquantized7.9B131K
Gemma 4 E4B IT OBLITERATED8.0B131K
Gemma4 E4B MiniFantasy V18.0B131K
Gemma 4 E4B IT8.0B131K
Gemma 4 E4B IT Ultra Uncensored Heretic8.0B131K
Gemma 4 E4B IT Uncensored8.0B131K
Gemma 4 E4B Luchador8.0B131K
Gemma 4 12B IT Heretic12.0B131K
Gemma 4 12B IT12.0B262K
Gemma 4 12B IT Qat Q4 0 Unquantized12.0B262K
Gemma 4 12B Coder Fable5 Composer2.5 V112.0B262K
Gemma 4 12B OBLITERATED12.0B131K
Gemma 4 12B IT AEON Abliterated K4 BF1612.0B262K
Gemma 4 12B12.0B262K
Gemma 4 12B Agentic Fable5 Composer2.5 v2 3.5x Tau212.0B262K
Gemma 4 12B IT Abliterated Uncensored12.0B131K
Gemma 4 12B IT Assistant12B262K
Gemma4 12B Mtp Assistant12B
Gemma 4 19B19.0B262K
Gemma 4 26B A4B IT Ultra Uncensored Heretic25.8B262K
Gemma 4 26B A4B IT Uncensored25.8B262K
Gemma 4 26B A4B IT Uncensored Heretic25.8B262K
Gemma 4 26B A4B IT Assistant26B262K
Gemma 4 26B A4B IT DFlash26B262K
Gemma 4 26B A4B IT26.5B262K
Gemma 4 26B A4B IT Qat Q4 0 Unquantized26.5B262K
Gemma 4 26B A4B StyleTune v226.5B262K
Gemma 4 31B IT Qat Q4 0 Unquantized Assistant31B131K
Gemma 4 31B IT Speculator.eagle331B
Gemma 4 31B IT DFlash31B262K
Gemma 4 31B IT Control Vectors31B
Gemma 4 31B IT Uncensored Heretic31.3B262K
Gemma 4 31B IT Heretic31.3B262K
Gemma 4 31B IT32.7B262K
Gemma 4 31B IT Qat Q4 0 Unquantized32.7B262K
Gemma 4 31B IT Uncensored32.7B262K
Gemma 4 31B StyleTune32.7B262K
ExtGemma4 40 5B39.5B262K

Frequently Asked Questions

How much VRAM do I need to run a Gemma 4 model?
The smallest Gemma 4 model, Gemma 4 E2B IT Qat Mobile Transformers, runs from 1.4 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
Which Gemma 4 models can I run on a 16 GB GPU?
41 of 47 Gemma 4 models fit in 16 GB of VRAM at some quantization, including Gemma 4 26B A4B IT, Gemma 4 31B IT, Gemma 4 12B IT.
What is the most popular Gemma 4 model to run locally?
Gemma 4 26B A4B IT is the most downloaded Gemma 4 model in local-friendly quantized formats. It runs from 8.0 GB of VRAM.
Gemma 4 Models — VRAM & Hardware Requirements | llmrun