◆ Local AI — Your private ChatGPT, free, on your own machine, in an hour · $24 · or all kits $49 →

Price tracking · recorded October 8, 2026

Graphics card prices for theLocal AI

18 graphics cards tracked twice a week at Materiel.net, compared on what matters for running an LLM: video memory, bandwidth, and price per GB.

The October 8 benchmarks

Report as of Thursday, October 8, 2026 · next report: Monday, October 12.

Cheapest GB of VRAM€37/GBIntel Arc B580 · 12 GB at €449.95
More speed per dollar7.8 tok/sIntel Arc B580 · for €100, on a 14B
24 GB and more at the best price1 199,95 €Intel Arc Pro B60 · 24 GB
32 GB at the best price1 799,95 €Intel Arc Pro B65 · 32 GB

All tracked cards

Lowest price at Materiel.net among cards in stock or shipping within 7 days, across all brands. Click a card's name to see which models it can run.

Sort by:
CardMateriel.net priceMemoryPrice per GBEstimated speed · bandwidthSince the previous updateBuy
Intel Arc B580Intel · 1 card in stock 449,95 € 12 GB 37 € ≈ 35 tok/s456 GB/s · 7.8 per €100 ▲ 20 € (+5 %)
Radeon RX 9060 XT 16 GBAMD · 9 cards in stock 819,95 € 16 GB 51 € ≈ 25 tok/s322 GB/s · 3.0 per €100 ▲ 20 € (+3 %)
RTX 5060 Ti 16 GBNVIDIA · 12 cards in stock 869,95 € 16 GB 54 € ≈ 35 tok/s448 GB/s · 4.0 per €100 ▲ 50 € (+6 %)
Radeon RX 9070 GREAMD · 2 cards in stock 899,95 € 12 GB 75 € ≈ 34 tok/s432 GB/s · 3.8 per €100 ▲ 200 € (+29 %)
RTX 5070NVIDIA · 7 cards in stock 949,95 € 12 GB 79 € ≈ 52 tok/s672 GB/s · 5.5 per €100 stable
Radeon RX 9070AMD · 5 cards in stock 999,95 € 16 GB 62 € ≈ 50 tok/s644 GB/s · 5.0 per €100 stable
Radeon RX 9070 XTAMD · 10 cards in stock 1 149,95 € 16 GB 72 € ≈ 50 tok/s644 GB/s · 4.3 per €100 ▲ 80 € (+7 %)
Intel Arc Pro B60Intel · 1 card in stock 1 199,95 € 24 GB 50 € ≈ 35 tok/s456 GB/s · 2,9 per 100 € ▲ 400 € (+50 %)
RTX 5070 TiNVIDIA · 15 cards in stock 1 479,95 € 16 GB 92 € ≈ 70 tok/s896 GB/s · 4.7 per €100 ▲ 80 € (+6 %)
Intel Arc Pro B65Intel · 1 card in stock 1 799,95 € 32 GB 56 € ≈ 47 tok/s608 GB/s · 2.6 per €100 ▲ 300 € (+20 %)
RTX 5080NVIDIA · 18 cards in stock 1 849,95 € 16 GB 116 € ≈ 75 tok/s960 GB/s · 4.1 per €100 stable
Intel Arc Pro B70Intel · 1 card in stock 2 299,95 € 32 GB 72 € ≈ 47 tok/s608 GB/s · 2.0 per €100 ▲ 400 € (+21 %)
Radeon AI PRO R9700AMD · 2 cards in stock 2 299,95 € 32 GB 72 € ≈ 50 tok/s640 GB/s · 2.2 per €100 ▲ 100 € (+5 %)
Radeon Pro W7800AMD · 1 card in stock 2 599,95 € 32 GB 81 € ≈ 45 tok/s576 GB/s · 1.7 per €100 stable
RTX 5090NVIDIA · 10 cards in stock 6 699,95 € 32 GB 209 € ≈ 139 tok/s1 792 GB/s · 2.1 per €100 ▼ 300 € (−4 %)
RTX PRO 4000 BlackwellNVIDIA · none in stock — 24 GB — ≈ 52 tok/s672 GB/s —
RTX PRO 5000 Blackwell 48 GBNVIDIA · none in stock — 48 GB — ≈ 105 tok/s1,344 GB/s —
RTX PRO 5000 Blackwell 72 GBNVIDIA · none in stock — 72 GB — ≈ 105 tok/s1,344 GB/s —

Estimated speed: about 70% of the theoretical ceiling (bandwidth ÷ model size) for a 14-billion-parameter model in Q4. The price column is from Materiel.net; the Amazon button displays its own current price. Affiliate links — commission paid by the merchant, at no extra cost to you; the table is calculated automatically.

What each amount of VRAM enables

The largest and newest models in the catalog that fit in Q4 with an 8K context, and the least expensive card offering at least that much memory according to the latest data. For an exact calculation, use the VRAM calculator or the configurator.

VRAMModels that fit (Q4)Cheapest card
12 GBNemotron 3 Super 12B, Gemma 4 12BIntel Arc B580 449,95 €
16 GBDevstral Small 2 24B, LFM2 24BRadeon RX 9060 XT 16 GB 819,95 €
24 GBQwen 3.6 35B-A3B, Laguna XS.2Intel Arc Pro B60 1 199,95 €
32 GBQwen 3.6 35B-A3B, Laguna XS.2Intel Arc Pro B65 1 799,95 €
48 GBNemotron 3 Puzzle 75B-A9B, LongCat Flash Lite Sparse 69B-A3B—
72 GBLaguna S 2.1, Qwen3-Coder-Next 80B-A3B—

Price trends

CardSep 28Oct. 1Oct 5Oct 8
RTX 50905 699,95 €6 391,95 €6 999,95 €6 699,95 €
RTX 50801 849,95 €1 738,95 €1 849,95 €1 849,95 €
RTX 5070 Ti1 399,95 €1 315,95 €1 399,95 €1 479,95 €
RTX 5070949,95 €911,75 €949,95 €949,95 €
RTX 5060 Ti 16 GB819,95 €770,75 €819,95 €869,95 €
Radeon RX 9070 XT1 069,95 €1 005,75 €1 069,95 €1 149,95 €
Radeon RX 9070999,95 €939,95 €999,95 €999,95 €
Radeon RX 9070 GRE699,95 €657,95 €699,95 €899,95 €
Radeon RX 9060 XT 16 GB799,95 €751,95 €799,95 €819,95 €
Intel Arc B580429,95 €404,15 €429,95 €449,95 €
Intel Arc Pro B60799,95 €751,95 €799,95 €1 199,95 €
Intel Arc Pro B651 499,95 €1 409,95 €1 499,95 €1 799,95 €
Intel Arc Pro B701 899,95 €1 785,95 €1 899,95 €2 299,95 €
Radeon AI PRO R97002 199,95 €2 067,95 €2 199,95 €2 299,95 €
Radeon Pro W78002 599,95 €2 443,95 €2 599,95 €2 599,95 €
RTX PRO 4000 Blackwell3 768,95 €3 542,81 €3 768,95 €—
RTX PRO 5000 Blackwell 48 GB9 899,95 €9 305,95 €——
RTX PRO 5000 Blackwell 72 GB10 439,95 €———

Amazon prices aren’t archived: Amazon only allows recent prices to be displayed. On Mac, unified memory changes the game: see which LLM for my Mac.

Frequently asked questions

Which graphics card should you buy to run an LLM locally?

Look at video memory (VRAM) first: the model must fit there entirely, or it spills into system memory and speed collapses. In Q4, 12 GB is enough for models up to about 14 billion parameters, 16 GB up to about 24 billion, 24 to 32 GB for models with 30 to 35 billion, and 48 GB up to about 70 billion. With equal memory, bandwidth is the deciding factor: it determines generation speed.

NVIDIA, AMD, or Intel for local AI?

NVIDIA remains the simplest choice: CUDA is supported by all the tools (Ollama, LM Studio, llama.cpp, vLLM). Radeon RX 9000 and AI PRO R9700 are supported by Ollama (ROCm 7), LM Studio and llama.cpp, sometimes with a bit more configuration. Intel Arc cards mainly use Vulkan or SYCL (llama.cpp, LM Studio) and OpenVINO: they often offer the cheapest VRAM, but the least mature ecosystem.

Is the RTX 5090 worth its price for AI?

This is the fastest consumer card (1,792 GB/s) with 32 GB, but it cost €6,699.95 in the latest check, compared with €2,349 at launch. For 32 GB, a Radeon AI PRO R9700 (€2,299.95) or an Intel Arc Pro B70 (€2,299.95) costs much less, at about 2.8 times lower speed. For very large models, a mini-PC with 128 GB of unified memory remains the alternative.

Where do these prices come from, and when are they updated?

Automatically collected every Monday and Thursday from our partner Materiel.net's product feed. For each chip, we select the least expensive card in stock or shipping within 7 days, across all brands. Amazon prices are not archived: Amazon only allows recent prices to be displayed, so they appear live on the buttons. The links are affiliate links: a purchase earns us a commission at no extra cost to you and does not influence the ranking, which is calculated automatically.

How is speed estimated?

To generate each token, the card rereads the entire model in memory: maximum speed is roughly bandwidth divided by model size. We use 70% of that figure for a 14-billion-parameter model in Q4 (about 9 GB). Actual measurements vary by software and drivers: this is an order-of-magnitude estimate for comparing cards, not a promise.

More free tools