GPU comparison
Compared on the specs that decide a purchase: memory, bandwidth, power, and which workloads each card actually fits — gaming, creative, or local models when that is the job.
Specifications verified August 10, 2026.
NVIDIA
GeForce RTX 4090
24GB VRAM · 1008 GB/s
NVIDIA
GeForce RTX 5090
32GB VRAM · 1792 GB/s
As an Amazon Associate I earn from qualifying purchases. Amazon UK links may earn us a commission; this does not change the products we include. Affiliate disclosure.
| Specification | GeForce RTX 4090 | GeForce RTX 5090 |
|---|---|---|
| Vendor | NVIDIA | NVIDIA |
| Memory | 24GB | Winner: 32GB |
| Usable for a model | 24GB | Winner: 32GB |
| Bandwidth | 1008 GB/s | Winner: 1792 GB/s |
| Segment | consumer | consumer |
| Released | October 12, 2022 | January 30, 2025 |
| Launch MSRP | ~$1,599~£1,259 | ~$1,999~£1,574 |
| UK street price | Search Amazon UK(paid link) (opens in new tab) | Search Amazon UK(paid link) (opens in new tab) |
Local models
Computed from published VRAM and the memory each model needs at Q4, including runtime overhead.
| Model | GeForce RTX 4090 | GeForce RTX 5090 |
|---|---|---|
| Qwen 3 4B4B parameters | Fits | Fits |
| Llama 3.1 8B Instruct8B parameters | Fits | Fits |
| Qwen 3 14B14B parameters | Fits | Fits |
| Gemma 3 27B27B parameters | Fits | Fits |
| Qwen 3 30B-A3B30B parameters | Fits | Fits |
| Qwen 3 32B32B parameters | Fits | Fits |
| Llama 3.3 70B70B parameters | Does not fit | Does not fit |
| Muse Glimmer 30B30B parameters | Fits | Fits |
| Mixtral 8x22B141B parameters | Does not fit | Does not fit |
| Qwen 3 235B-A22B235B parameters | Does not fit | Does not fit |
| Qwen 3 Coder 480B-A35B480B parameters | Does not fit | Does not fit |
| DeepSeek V3.2671B parameters | Does not fit | Does not fit |
Common questions
Answered from the verified figures on this page rather than general guidance.
GeForce RTX 4090 launched lower at ~$1599 against ~$1999 for GeForce RTX 5090. Launch MSRP is an anchor, not a live price.
GeForce RTX 5090 has more memory bandwidth — 1792 GB/s against 1008 GB/s. That usually helps high-refresh 1440p and 4K once VRAM is sufficient for your settings. Still verify with dated independent tests for the games you play.
GeForce RTX 5090 — 32GB usable against 24GB on GeForce RTX 4090. Extra memory helps texture mods, video timelines, and 3D scenes before it helps esports frame rates.
Both hold the same models from this list at Q4, but GeForce RTX 5090 has more memory — 32GB against 24GB. That headroom buys longer context or a higher quantisation rather than a bigger model.
GeForce RTX 5090, at 1792 GB/s against 1008 GB/s — roughly 1.8× the memory bandwidth. Bandwidth is the practical ceiling on token throughput once a model fits, but it only matters if the model fits in the first place.
Neither. A 70B model needs about 42GB at Q4 including overhead, above both cards. The largest that fits is Qwen 3 32B on GeForce RTX 4090 and Qwen 3 32B on GeForce RTX 5090.
If one card holds a model the other cannot, that capacity difference usually outweighs a bandwidth advantage. A model that spills to system memory can be substantially slower, but the actual impact depends on the backend, transfer path, context, and workload.