Baichuan Models — Hardware Requirements

5 Baichuan models from baichuan-inc and the community, from the smallest that runs in 3.3 GB of VRAM up to 13B parameters. Every row links to full quantization tables and GPU compatibility.

All Baichuan Models by Size

ModelParamsContext
Baichuan 7B7B4K
Baichuan2 7B Base7B4K
Baichuan2 13B Chat13B—
Baichuan 13B Base13B—
Baichuan2 13B Base13B—

Frequently Asked Questions

How much VRAM do I need to run a Baichuan model?
The smallest Baichuan model, Baichuan2 7B Base, runs from 3.3 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
Which Baichuan models can I run on a 16 GB GPU?
5 of 5 Baichuan models fit in 16 GB of VRAM at some quantization, including Baichuan2 13B Chat, Baichuan 13B Base, Baichuan2 13B Base.
What is the most popular Baichuan model to run locally?
Baichuan2 13B Chat is the most downloaded Baichuan model in local-friendly quantized formats. It runs from 3.9 GB of VRAM.
Baichuan Models — VRAM & Hardware Requirements | llmrun