NVIDIA GeForce RTX 4090 vs NVIDIA GeForce RTX 5090, AI & Machine Learning Comparison

NVIDIA GeForce RTX 4090
NVIDIA GeForce RTX 4090
vs
NVIDIA GeForce RTX 5090
NVIDIA GeForce RTX 5090

NVIDIA GeForce RTX 5090 wins 59 of 61 benchmarks, averaging 43.9% faster.

Both cards were measured first-party on our bench, same suite, same test rig.

What the numbers say

The gap is widest in Wan 2.2 TI2V-5B (image to video), where NVIDIA GeForce RTX 5090 leads by 104% (1.122 vs 2.284 clips/min); the closest fight is Stable Diffusion 2.1 (8% apart); VRAM decides part of this one: NVIDIA GeForce RTX 5090 runs 69 of our 12 AI workloads while the other card runs 63, models that don't fit score zero.

Benchmark results head-to-head

BenchmarkNVIDIA GeForce RTX 4090NVIDIA GeForce RTX 5090Difference
Qwen3 0.6B tok/s770.13868.61-11%
Llama 3.2 1B tok/s752.521060.14-29%
MiniCPM5 2B tok/s372.72507.31-27%
LFM2.5 2.6B tok/s385.05566.57-32%
Granite 4.1 3B tok/s276.62375.37-26%
Agents-A1-4B tok/s219.4328.96-33%
Nemotron 3 Nano 4B tok/s258.12410.94-37%
Qwen3 4B tok/s260.54375.6-31%
Spark-X2.5-4B tok/s245.97352.29-30%
DeepSeek Coder 7B Instruct v1.5 tok/s188.59291.38-35%
OLMo 3 7B Instruct tok/s170.75266.68-36%
OLMo 3 7B Think tok/s170.69266.72-36%
Qwen2-7B-Instruct tok/s177.4277.64-36%
Qwen2.5-7B tok/s183.66284.6-35%
Qwen2.5-Coder 7B tok/s183.73284.59-35%
Apertus-8B-Instruct tok/s163.09257.69-37%
Llama 3 8B tok/s167.69259.18-35%
Llama 3.1 8B tok/s171.29268.14-36%
Qwen3 8B tok/s164.32243.91-33%
Nemotron Nano 9B v2 tok/s121.89202.68-40%
Ornith 1.5 9B tok/s144.04225.79-36%
Gemma 4 12B tok/s103.57151.54-32%
Qwen2.5-Coder 14B tok/s95.12149.75-36%
Qwen3 14B tok/s96.37143.11-33%
gpt-oss-20b tok/s286.23415.12-31%
Gemma 4 26B A4B tok/s188.23264.69-29%
Qwen3.6 27B tok/s49.0276.6-36%
Qwen3.8 27B tok/s48.0175.74-37%
Qwen3 30B A3B tok/s259.61344.67-25%
Qwen3 30B A3B Instruct 2507 tok/s263.17352.33-25%
Qwen3-Coder 30B A3B tok/s271.04365.84-26%
Gemma 4 31B tok/s45.0871.42-37%
Qwen3 32B tok/s44.2871.15-38%
Llama 3.3 70B tok/s00n/a
Stable Diffusion 1.5 images/min80.14102.36-22%
Stable Diffusion 2.1 images/min60.5865.53-8%
SDXL Turbo images/min652.11902.04-28%
SSD-1B images/min15.7720.49-23%
Z-Image Turbo images/min7.2410.725-32%
Sana 1.6B images/min41.4751.52-20%
Stable Diffusion XL images/min16.2821.08-23%
DreamShaper XL Lightning images/min91.52122.85-26%
DreamShaper XL Turbo images/min54.3570.72-23%
Playground v2.5 images/min9.9613.28-25%
PixArt-Sigma XL images/min24.2730.01-19%
FLUX.2 klein 4B images/min42.3859.33-29%
Kolors images/min9.8813.53-27%
Z-Image images/min1.171.72-32%
AuraFlow v0.3 images/min3.024.35-31%
FLUX.1 dev images/min02.036n/a
Qwen-Image-Edit images/min00n/a
Wan 2.1 1.3B frames/s0.6730.879-23%
CogVideoX-2B frames/s0.5170.7-26%
CogVideoX-5B frames/s0.170.259-34%
LTX-Video (distilled) frames/s6.69610.449-36%
Stable Video Diffusion clips/min2.2352.836-21%
Wan 2.2 5B (720p) frames/s0.430.588-27%
LTX-Video (image to video) clips/min4.1886.152-32%
Wan 2.2 TI2V-5B (image to video) clips/min1.1222.284-51%
Stable Video Diffusion XT clips/min1.2121.344-10%
CogVideoX-5B I2V clips/min0.3280.447-27%

Whole-job comparison

How long each card takes to finish a complete pipeline, not just one model. NVIDIA GeForce RTX 4090 is faster on 3 of 5; NVIDIA GeForce RTX 5090 on 2.

WorkflowNVIDIA GeForce RTX 4090NVIDIA GeForce RTX 5090DifferenceCost per run
24-frame storyboard3.8 min12.5 minNVIDIA GeForce RTX 4090 3.24x faster$0.022 vs $0.096
60-second AI short film5.8 min9.3 minNVIDIA GeForce RTX 4090 1.60x faster$0.033 vs $0.072
Full codebase review10.5 min6.7 minNVIDIA GeForce RTX 5090 1.57x faster$0.060 vs $0.052
Animate a batch of images17.8 min8.8 minNVIDIA GeForce RTX 5090 2.04x faster$0.101 vs $0.068
Short social clips21.5 min24.8 minNVIDIA GeForce RTX 4090 1.15x faster$0.122 vs $0.192

Renting by the hour, NVIDIA GeForce RTX 4090 finishes 3 of 5 cheaper. The quicker card is not automatically the cheaper way to get the work done.

Cost to rent

CardPer hour
NVIDIA GeForce RTX 4090$0.340
NVIDIA GeForce RTX 5090$0.463

NVIDIA GeForce RTX 4090 is 1.36x cheaper per hour. Cheapest on-demand rate we see across RunPod and Vast.

Specifications compared

NVIDIA GeForce RTX 4090NVIDIA GeForce RTX 5090
VRAM24GB32GB
Transistors76,300M92,200M
Die size608.4 mm²750 mm²
Process node4 nm4 nm
Transistor density125.4 M/mm²122.9 M/mm²
ArchitectureAda LovelaceBlackwell (GB202)
Memory bandwidth1008 GB/s1792 GB/s
Boost clock2,520 MHz2,407 MHz
TDP450 W575 W
Launch MSRP$1,599$1,999
Release2022-10-122025-01-30

FAQ

Which is better for ai & machine learning: NVIDIA GeForce RTX 4090 or NVIDIA GeForce RTX 5090?
NVIDIA GeForce RTX 5090 performs better for ai & machine learning, winning 59 of 61 benchmarks in our suite with an average 43.9% advantage.
What are the main hardware differences between NVIDIA GeForce RTX 4090 and NVIDIA GeForce RTX 5090?
NVIDIA GeForce RTX 4090 has 24GB VRAM and a 450W TDP, while NVIDIA GeForce RTX 5090 has 32GB VRAM and a 575W TDP.
Does VRAM matter more than speed between NVIDIA GeForce RTX 4090 and NVIDIA GeForce RTX 5090?
For AI, yes, NVIDIA GeForce RTX 5090 fits 2 more of our 12 workloads than NVIDIA GeForce RTX 4090. A model that exceeds VRAM doesn't run slower, it doesn't run at all, so the card that fits the model wins that workload outright.
Where is the biggest performance difference between NVIDIA GeForce RTX 4090 and NVIDIA GeForce RTX 5090?
Wan 2.2 TI2V-5B (image to video): NVIDIA GeForce RTX 5090 leads by roughly 104% (1.122 vs 2.284 clips/min) in our testing.

NVIDIA GeForce RTX 4090 full review · NVIDIA GeForce RTX 5090 full review · All AI & Machine Learning rankings