Home›Catalog›RTX 5090 vs RTX 5080

RTX 5090 vs. RTX 5080: which GPU for local AI?

Focused comparison Local LLM : we're looking at what really matters for running models—VRAM above all, then generation speed and price. Purchase verdict at the end.

Updated on 09/10/2026

Comparison at a glance

CriterionRTX 5090RTX 5080
Memory32 GB VRAM16 GB VRAM
BrandNVIDIANVIDIA
Compatible models (Q4)182 models in the catalog143 models in the catalog
Largest model (Q4)Jamba 1.5 Mini (52B)Qwen 3.8 27B Obliterated (28B)

Which one should you choose to run an LLM?

For large models (Q4 30B+ or a 70B), prioritize the RTX 5090 and its 32 GB—the VRAM is the No. 1 factor for LLMs. The RTX 5080 remains relevant if your budget is tighter or you're targeting models ≤ 16 GB (comfortable at 7–24B).

Golden rule: with the same budget, always choose the GPU with the most VRAM. A model that exceeds the VRAM limit collapses in speed (CPU offloading). VRAM determines the size of the model; computing power mainly dictates the Speed.

Where to buy (up-to-date pricing and stock)

RTX 5090
AmazonSee the price →

Alternative purchase: GMKtec EVO-X2 64 GB / 1 TB (Ryzen AI Max+ 395). A mini PC is a complete machine, not a replacement board; check memory and software compatibility.

Affiliate links — BestLLMfor may earn a commission from purchases at no extra cost to you, which does not influence this independently determined comparison. As an Amazon Associate, BestLLMfor earns from qualifying purchases.

Frequently asked questions

RTX 5090 or RTX 5080 for running an LLM locally?

For local use, the VRAM premium. The RTX 5090 (32 GB) runs larger models than the RTX 5080 (16 GB).

What is the largest model the RTX 5090 can run?

Up to Jamba 1.5 Mini (52B) in Q4 quantization, with its 32 GB.

Should you buy used?

For LLMs, a used GPU with more VRAM often beats a newer card with less memory. Check its condition and warranty.

Go further