Arc Pro B60: 24GB VRAM — What Local LLMs Can It Run?
Check prices
The Arc Pro B60 has 24GB of GDDR6 memory with 456 GB/s of bandwidth. No used-price data yet. At Q4 quantization it runs models up to 32B parameters at roughly 11 tokens/s.
Full Arc Pro B60 specifications
What LLMs can the Arc Pro B60 run?
One reference model per family (its main version)| Model | Q4_K_M | Q8_0 | FP16 | ||||||
|---|---|---|---|---|---|---|---|---|---|
| Fits | Speed t/s | Context | Fits | Speed t/s | Context | Fits | Speed t/s | Context | |
| DeepSeekDeepSeek R1 Qwen3 8B | ✓ | 38 | 128K | ✓ | 21.8 | 32K | ✓ | 12 | 32K |
| GemmaGemma 4 31B | ✗ | — | — | ✗ | — | — | ✗ | — | — |
| GLMGLM 4.7 Flash | ✓ | 75 | 32K | Offload | 19.9 | — | Offload | 11.1 | — |
| gpt-ossgpt-oss 20B | ✓ | 67.1 | 128K | ✓ | 42.5 | 8K | Offload | 8.6 | — |
| KimiKimi Linear 48B A3B | Offload | 39 | — | Offload | 22.7 | — | Offload | 12.6 | — |
| LlamaLlama 3.1 8B | ✓ | 39.2 | 128K | ✓ | 22.4 | 32K | ✓ | 12.3 | 32K |
| MistralMistral 7B v0.3 | ✓ | 42.8 | 32K | ✓ | 24.6 | 32K | ✓ | 13.5 | 32K |
| QwenQwen3 8B | ✓ | 38 | 32K | ✓ | 21.8 | 32K | ✓ | 12 | 32K |
Fits when weights + KV cache + runtime overhead is at or below usable memory, single card. "Offload" on MoE models means it runs with expert weights in system RAM (only attention layers and KV cache stay in VRAM); that speed assumes 70 GB/s RAM bandwidth. The context column is the largest tier that fits.
Arc Pro B60 used price history
Not enough history to draw a chart yet — price tracking just started.
Prices by platform
Buy links may carry affiliate tags; they do not change the price you pay.
FAQ
How much VRAM does the Arc Pro B60 have?
The Arc Pro B60 has 24GB of GDDR6 memory with 456 GB/s of bandwidth.
Can the Arc Pro B60 run a 70B model?
Not on a single card. At Q4_K_M with an 8K context a 70B dense model needs about 43GB, more than the Arc Pro B60's 24GB.
How much does a used Arc Pro B60 cost?
No used-price data yet; it will appear here once the feed is live.