GPUs tracked39Price history3 dayssince Sep 3, 2026Price points (24h)11,753 pointsBiggest 24h dropBiggest 24h riseData updated Sep 5, 2026

What Do You Need to Run Mistral Locally? VRAM by Version

Mistral currently has 5 versions you can deploy locally. The smallest, Mistral 7B v0.3, needs 7GB of VRAM at Q4; the largest, Mistral Small 3.2 24B, needs 16GB. The cheapest device that gets one running is the GeForce RTX 2080 Ti, with no used-price data yet.

VRAM needed for each Mistral version

5 versions
VersionReleasedContext
Mistral 7B v0.3Dense7.3BMay 22, 202432K7GeForce RTX 2080 Ti · 11 GB2,648,636
Ministral 3 8BDense8.9BOct 31, 2025256K8GeForce RTX 2080 Ti · 11 GB142,610
Mistral NemoDense12.3BJul 17, 2024128K10GeForce RTX 2080 Ti · 11 GB358,821
Ministral 3 14BDense14BOct 31, 2025256K11GeForce RTX 2080 Ti · 11 GB266,145
Mistral Small 3.2 24BDense24BJun 19, 2025128K16GeForce RTX 4080 · 16 GB126,611

Minimum VRAM = Q4_K_M weights + KV cache for an 8K context + 1GB runtime overhead, rounded up. The cheapest device is the first one that fits when sorting by median used price on the current price basis.

About Mistral

Mistral is published by Mistral AI under Apache-2.0. This page covers 5 versions ranging from 7.3B to 24B parameters; structure parameters and quantized sizes are synced weekly from Hugging Face.

VendorMistral AI
LicenseApache-2.0
Official sitehttps://mistral.ai/
Hugging Face orgmistralai