Mixtral Models — Hardware Requirements

7 Mixtral models from Mistral AI and the community, from the smallest that runs in 0.4 GB of VRAM up to 140.6B parameters. Every row links to full quantization tables and GPU compatibility.

All Mixtral Models by Size

ModelParamsContext
Tiny Mixtral247M131K
Mixtral 8x7B Instruct v0.146.7B33K
Mixtral 8x7B v0.146.7B33K
Nous Hermes 2 Mixtral 8x7B DPO46.7B33K
Mixtral 34Bx2 MoE 60B60.8B200K
Mixtral 8x22B v0.1140.6B66K
Mixtral 8x22B Instruct v0.1140.6B66K

How Mixtral Compares — Benchmark Rating

Mixtral 8x22B Instruct v0.1 is the highest-rated Mixtral model with an overall benchmark rating of 24.4/100 — #68 among 78 open models. The top proprietary model, Claude Fable 5 (high), scores 87.8. Click a model to see its full benchmark breakdown.

Claude Fable 5 (high) · proprietary87.8
Claude Opus 4.7 · proprietary86.4
Claude Fable 5 Max · proprietary82.4
GPT 6 Astra (high) · proprietary82.4
GPT 6 Astra Max · proprietary82.4
GLM 565.9
Composite of normalized public benchmark scores (methodology) · ■ Mixtral · ■ other models

Frequently Asked Questions

How much VRAM do I need to run a Mixtral model?
The smallest Mixtral model, Tiny Mixtral, runs from 0.4 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
Which Mixtral models can I run on a 16 GB GPU?
1 of 7 Mixtral models fit in 16 GB of VRAM at some quantization, including Tiny Mixtral.
What is the most popular Mixtral model to run locally?
Mixtral 8x7B Instruct v0.1 is the most downloaded Mixtral model in local-friendly quantized formats. It runs from 20.4 GB of VRAM.
How do Mixtral models score on benchmarks?
Mixtral 8x22B Instruct v0.1 leads the family with an overall benchmark rating of 24.4/100, ranking #68 among 78 open models, while the top proprietary model, Claude Fable 5 (high), scores 87.8. See the comparison chart above for the full standings.
Mixtral Models — VRAM & Hardware Requirements | llmrun