Qwen 3.8 Models — Hardware Requirements

10 Qwen 3.8 models from Alibaba and the community, from the smallest that runs in 11.8 GB of VRAM up to 2446.2B parameters. Every row links to full quantization tables and GPU compatibility.

All Qwen 3.8 Models by Size

ModelParamsContext
Qwen3.8 2B2.3B262K
Qwen3.8 4B4.7B262K
Qwen3.8 4B Distill4.7B262K
Qwen3.8 27B Escha W26.3B262K
Qwen3.8 27B EXL3 3.5bpw7.7B262K
Qwen3.8 9B Distill Uncensored Heretic9.4B262K
Qwen3.8 9B9.7B262K
Qwen3.8 9B Distill9.7B262K
Qwen3.8 Flash Next VQ 2.1bpw24.6B262K
Qwen3.8 27B Abliterated MTPLX Optimized Speed26.9B262K
Qwen3.8 27B Obliterated E0326.9B262K
Qwen3.8 Whittle MoE 27B A17.8B26.9B262K
Qwen3.8 27B DSpark27B262K
Qwen3.8 27B DFlash227B262K
Qwen3.8 27B DFlash227B262K
Qwen3.8 27B EfficientThink Uncensored K3 Opus5 Grok4.6 GPT5.6Sol SFT SimPO DFlash227B
Qwen3.8 27B ASCII Condensed27B
Qwen3.8 27B DSpark Agentic27B262K
Qwen3.8 27B MTPLX Optimized Speed27.4B262K
Qwen3.8 27B MTPLX Optimized Quality27.4B262K
Qwen3.8 27B Uncensored27.4B262K
Ektome Qwen3.8 27B PristinelyUncensored27.4B262K
Grug V1.1 Qwen 3.8 27B27.4B262K
Qwen3.8 27B27.8B262K
Huihui Qwen3.8 27B Abliterated27.8B262K
Qwen3.8 27B AEON ULTIMATE UNCENSORED BF1627.8B262K
Qwen3.8 27B TURBO Fable Cold Fusion 735 882 Heretic Uncensored NM DAU27.8B262K
Qwen3.8 27B OBLITERATED27.8B262K
Qwen3.8 27B Uncensored27.8B
Qwen3.8 27B ABLITERATED BF1627.8B262K
Qwen3.8 27B Cold Fusion GAIN V1.127.8B262K
Qwen3.8 Distill 35B A3B Coder Abliterated35B262K
Qwen3.8 Flash Coder 85gb BF1642.6B262K
Qwen3.8 Flash Next MTPLX Optimized Speed126.2B262K
Qwen3.8 Flash Next MTPLX Bare Speed126.2B262K
Qwen3.8 Flash Next180.0B262K
Qwen3.8 Flash Next Uncensored180.0B
Qwen3.8 2.4T A95B2446.2B262K

How Qwen 3.8 Compares — Benchmark Rating

Qwen3.8 Flash Next is the highest-rated Qwen 3.8 model with an overall benchmark rating of 81.9/100 — #3 among 123 open models. The top proprietary model, GPT 5.5, scores 89.2. Click a model to see its full benchmark breakdown.

GPT 5.5 · proprietary89.2
Muse Spark 1.3 · proprietary88.9
Claude Opus 4.7 · proprietary87.4
Claude Sonnet 5 · proprietary87.4
Gemini 3.1 Pro Preview (high) · proprietary87.2
Composite of normalized public benchmark scores (methodology) · Qwen 3.8 · other models

Frequently Asked Questions

How much VRAM do I need to run a Qwen 3.8 model?
The smallest Qwen 3.8 model, Qwen3.8 2B, runs from 1.4 GB of VRAM at an aggressive quantization. Larger family members need proportionally more — see the table above for every model.
Which Qwen 3.8 models can I run on a 16 GB GPU?
22 of 38 Qwen 3.8 models fit in 16 GB of VRAM at some quantization, including Qwen3.8 27B, Qwen3.8 Whittle MoE 27B A17.8B, Qwen3.8 27B DSpark.
What is the most popular Qwen 3.8 model to run locally?
Qwen3.8 27B is the most downloaded Qwen 3.8 model in local-friendly quantized formats. It runs from 12.6 GB of VRAM.
How do Qwen 3.8 models score on benchmarks?
Qwen3.8 Flash Next leads the family with an overall benchmark rating of 81.9/100, ranking #3 among 123 open models, while the top proprietary model, GPT 5.5, scores 89.2. See the comparison chart above for the full standings.