Ling Models — Hardware Requirements

3 Ling models from Inclusion AI and the community, from the smallest that runs in 2.8 GB of VRAM up to 127.5B parameters. Every row links to full quantization tables and GPU compatibility.

All Ling Models by Size

ModelParamsContext
Ling 3.0 Tiny7.9B131K
Ling 3.0 Flash VL124.8B131K
Ling 3.0 Flash127.5B262K
Ling 3.0 Flash Fin127.5B262K

Frequently Asked Questions

How much VRAM do I need to run a Ling model?
The smallest Ling model, Ling 3.0 Tiny, runs from 2.8 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
Which Ling models can I run on a 16 GB GPU?
1 of 4 Ling models fit in 16 GB of VRAM at some quantization, including Ling 3.0 Tiny.
What is the most popular Ling model to run locally?
Ling 3.0 Tiny is the most downloaded Ling model in local-friendly quantized formats. It runs from 2.8 GB of VRAM.