Gemma 4 Models — Hardware Requirements
25 Gemma 4 models from Google and the community, from the smallest that runs in 2 GB of VRAM up to 32.7B parameters. Every row links to full quantization tables and GPU compatibility.
All Gemma 4 Models by Size
Frequently Asked Questions
- How much VRAM do I need to run a Gemma 4 model?
- The smallest Gemma 4 model, Gemma 4 E2B IT Qat Mobile Transformers, runs from 1.4 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
- Which Gemma 4 models can I run on a 16 GB GPU?
- 41 of 47 Gemma 4 models fit in 16 GB of VRAM at some quantization, including Gemma 4 26B A4B IT, Gemma 4 31B IT, Gemma 4 12B IT.
- What is the most popular Gemma 4 model to run locally?
- Gemma 4 26B A4B IT is the most downloaded Gemma 4 model in local-friendly quantized formats. It runs from 8.0 GB of VRAM.