GPU comparison
Compared on the specs that decide a purchase: memory, bandwidth, power, and which workloads each card actually fits — gaming, creative, or local models when that is the job.
Specifications verified August 10, 2026.
NVIDIA
GeForce RTX 5080
16GB VRAM · 960 GB/s
NVIDIA
GeForce RTX 5090
32GB VRAM · 1792 GB/s
As an Amazon Associate I earn from qualifying purchases. Amazon UK links may earn us a commission; this does not change the products we include. Affiliate disclosure.
| Specification | GeForce RTX 5080 | GeForce RTX 5090 |
|---|---|---|
| Vendor | NVIDIA | NVIDIA |
| Memory | 16GB | Winner: 32GB |
| Usable for a model | 16GB | Winner: 32GB |
| Bandwidth | 960 GB/s | Winner: 1792 GB/s |
| Segment | consumer | consumer |
| Released | January 30, 2025 | January 30, 2025 |
| Launch MSRP | ~$999~£787 | ~$1,999~£1,574 |
| UK street price | Search Amazon UK(paid link) (opens in new tab) | Search Amazon UK(paid link) (opens in new tab) |
Local models
Computed from published VRAM and the memory each model needs at Q4, including runtime overhead.
| Model | GeForce RTX 5080 | GeForce RTX 5090 |
|---|---|---|
| Qwen 3 4B4B parameters | Fits | Fits |
| Llama 3.1 8B Instruct8B parameters | Fits | Fits |
| Qwen 3 14B14B parameters | Fits | Fits |
| Gemma 3 27B27B parameters | Does not fit | Fits |
| Qwen 3 30B-A3B30B parameters | Does not fit | Fits |
| Qwen 3 32B32B parameters | Does not fit | Fits |
| Llama 3.3 70B70B parameters | Does not fit | Does not fit |
| Muse Glimmer 30B30B parameters | Does not fit | Fits |
| Mixtral 8x22B141B parameters | Does not fit | Does not fit |
| Qwen 3 235B-A22B235B parameters | Does not fit | Does not fit |
| Qwen 3 Coder 480B-A35B480B parameters | Does not fit | Does not fit |
| DeepSeek V3.2671B parameters | Does not fit | Does not fit |
Common questions
Answered from the verified figures on this page rather than general guidance.
GeForce RTX 5080 launched lower at ~$999 against ~$1999 for GeForce RTX 5090. Launch MSRP is an anchor, not a live price.
GeForce RTX 5090 has more memory bandwidth — 1792 GB/s against 960 GB/s. That usually helps high-refresh 1440p and 4K once VRAM is sufficient for your settings. Still verify with dated independent tests for the games you play.
GeForce RTX 5090 — 32GB usable against 16GB on GeForce RTX 5080. Extra memory helps texture mods, video timelines, and 3D scenes before it helps esports frame rates.
GeForce RTX 5090. With 32GB usable against 16GB it additionally holds Gemma 3 27B, Qwen 3 30B-A3B, Qwen 3 32B, Muse Glimmer 30B at Q4.
GeForce RTX 5090, at 1792 GB/s against 960 GB/s — roughly 1.9× the memory bandwidth. Bandwidth is the practical ceiling on token throughput once a model fits, but it only matters if the model fits in the first place.
Neither. A 70B model needs about 42GB at Q4 including overhead, above both cards. The largest that fits is Qwen 3 14B on GeForce RTX 5080 and Qwen 3 32B on GeForce RTX 5090.
If one card holds a model the other cannot, that capacity difference usually outweighs a bandwidth advantage. A model that spills to system memory can be substantially slower, but the actual impact depends on the backend, transfer path, context, and workload.