Gemma Models — Hardware Requirements
14 Gemma models from Google and the community, from the smallest that runs in 0.3 GB of VRAM up to 27.0B parameters. Every row links to full quantization tables and GPU compatibility.
All Gemma Models by Size
| Model | Params | Runs from | Context | Publisher | Quant downloads |
|---|---|---|---|---|---|
| Functiongemma 270M IT | 268M | 0.6 GB | — | ||
| Functiongemma 270M Ft Mobile Actions | 270M | 0.6 GB | — | ||
| T5gemma B B Ul2 IT | 591M | 1.3 GB | — | ||
| Vaultgemma 1B | 1.0B | 2.3 GB | — | ||
| T5gemma L L Ul2 IT | 1.2B | 2.7 GB | — | ||
| Gemma 1.1 2B IT | 2.5B | 1.1 GB | — | ||
| Medgemma 1.5 4B IT | 4.3B | 1.3 GB | — | ||
| Gemma 7B IT | 8.5B | 18.8 GB | — | ||
| Gemma 7B | 8.5B | 18.8 GB | — | ||
| Codegemma 7B IT | 8.5B | 4.0 GB | — | ||
| Turkish Gemma 9B T1 | 9.2B | 4.8 GB | 8K | ||
| Turkish Gemma 9B v0.1 | 9.2B | 4.8 GB | 8K | ||
| Medgemma 27B Text IT | 27.0B | 59.4 GB | — |
Frequently Asked Questions
- How much VRAM do I need to run a Gemma model?
- The smallest Gemma model, Functiongemma 270M IT, runs from 0.6 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
- Which Gemma models can I run on a 16 GB GPU?
- 10 of 13 Gemma models fit in 16 GB of VRAM at some quantization, including Medgemma 1.5 4B IT, Gemma 1.1 2B IT, Functiongemma 270M IT.
- What is the most popular Gemma model to run locally?
- Medgemma 27B Text IT is the most downloaded Gemma model in local-friendly quantized formats. It runs from 59.4 GB of VRAM.