Gemma Models — Hardware Requirements

14 Gemma models from Google and the community, from the smallest that runs in 0.3 GB of VRAM up to 27.0B parameters. Every row links to full quantization tables and GPU compatibility.

All Gemma Models by Size

ModelParamsContext
Functiongemma 270M IT268M—
Functiongemma 270M Ft Mobile Actions270M—
T5gemma B B Ul2 IT591M—
Vaultgemma 1B1.0B—
T5gemma L L Ul2 IT1.2B—
Gemma 1.1 2B IT2.5B—
Medgemma 1.5 4B IT4.3B—
Gemma 7B IT8.5B—
Gemma 7B8.5B—
Codegemma 7B IT8.5B—
Turkish Gemma 9B T19.2B8K
Turkish Gemma 9B v0.19.2B8K
Medgemma 27B Text IT27.0B—

Frequently Asked Questions

How much VRAM do I need to run a Gemma model?
The smallest Gemma model, Functiongemma 270M IT, runs from 0.6 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
Which Gemma models can I run on a 16 GB GPU?
10 of 13 Gemma models fit in 16 GB of VRAM at some quantization, including Medgemma 1.5 4B IT, Gemma 1.1 2B IT, Functiongemma 270M IT.
What is the most popular Gemma model to run locally?
Medgemma 27B Text IT is the most downloaded Gemma model in local-friendly quantized formats. It runs from 59.4 GB of VRAM.