GeForce RTX 4090 vs 3090 Ti

We compared two discrete desktop gaming GPUs: the GeForce RTX 4090 with 128 pipelines and 16384 shaders against the 6 months older RTX 3090 Ti that utilizes 84 pipelines and 10752 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.

Review

General comparison of performance in games, applications, power efficiency, and other metrics
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score

Value for money (Beta)

We use average GPU prices from all manufacturers by default. Enter the current price of a specific card for better accuracy.
Currency
24 GB GDDR6X
VS
24 GB GDDR6X
18.4
(Bad)
Value Index
34.2
(Mediocre)

Key differences

Key distinctions and advantages of RTX 3090 Ti over RTX 4090
Reasons to consider the GeForce RTX 4090
  • Performs significantly better (up to 73%) in 3DMark Steel Nomad Lite
  • 2.1x higher maximum theoretical performance (82.6 vs 40 TFLOPS)
  • Manufactured using a more efficient 5 nm process technology
  • Shows 39% higher average frame rate in modern games at QHD resolution – 177 vs 127 FPS
  • Features 176 more tensor cores for effective ML and AI workloads
  • Supports Nvidia DLSS 3 technology
  • Achieves 36% more points in the GeekBench 6 Compute test (316K vs 232K)
  • Has 52% more shading units (16384 vs 10752)

Gaming Performance

Frame rate comparison across popular AAA titles at different resolutions

Games

FPS Table
Forza Horizon 5
275
199
193
155
200
136
126
99
The Witcher 3
292
268
191
101
214
185
130
71
Counter-Strike 2
339
238
188
90
286
213
154
85
Far Cry 6
203
179
170
104
160
151
133
83
Hogwarts Legacy
200
163
132
74
141
116
93
52
CoD: Modern Warfare III
266
263
219
151
186
174
139
90
Ghost of Tsushima
188
152
137
91
129
105
89
57
Cyberpunk 2077
226
208
143
66
158
140
98
44
Shadow of the Tomb Raider
302
284
222
118
254
237
183
101
Average FPS by Resolution
1080p High 255 192
1080p Ultra 217 162
1440p Ultra 177 127
4K Ultra 106 76
Margin of Error Low Low

Benchmarks

Graphics cards’ performance in recent benchmarking apps

3D Mark

Multiplatform graphics benchmark suite that directly correlates with performance in modern games
Steel Nomad Lite Score
RTX 4090 +73%
42169
24378
Time Spy 36299 21855
Solar Bay 187525 109239
Port Royal 26127 14890
Fire Strike 72432 51935
Wild Life Extreme 85218 48861
Night Raid 195813 156537
Sources: 3DMark [1], [2]

GeekBench 6 OpenCL

GPU test for computational tasks (image processing, photography, computer vision, and ML)
GB6 Compute Score
RTX 4090 +36%
316301
232083
Background Blur 318.9 img/sec 362.7 img/sec
Face Detection 222.8 img/sec 227.9 img/sec
Horizon Detection 14.2 Gpixels/sec 11.3 Gpixels/sec
Edge Detection 21 Gpixels/sec 15.9 Gpixels/sec
Gaussian Blur 23.9 Gpixels/sec 12.1 Gpixels/sec
Feature Matching 2.38 Gpixels/sec 1.56 Gpixels/sec
Stereo Matching 1950 Gpixels/sec 1080 Gpixels/sec
Particle Physics 48090.4 FPS 31024.6 FPS
API OpenCL OpenCL
Sources: Geekbench [3], [4]

Cinebench 2024 GPU

Hardware benchmark using Maxon's Cinema 4D rendering engine
Cinebench 2024 GPU
RTX 4090 +60%
32578
20322

Passmark Graphics

Videocard test that focuses on compute shaders, multi-texturing, tessellation, and other features
G3D Mark Score
RTX 4090 +30%
38042
29247
G2D Mark 1303 1226
DirectX 11 323 FPS 238 FPS
DirectX 12 151 FPS 121 FPS
GPU Compute 26022 Ops/s 16494 Ops/s
Sources: PassMark [5], [6] – 22248 & 3325 samples

Blender

Rendering performance test for 3D modeling
Blender GPU
RTX 4090 +93%
11682.55
6050.04
Sources: Blender [9], [10] – 8909 & 1861 samples

Recent User Tests

The latest benchmark tests that have been submitted by users
GeForce RTX 4090
DateBenchmarkResult
📘 2026-04-03 (afroman420IU)Cinebench 202432578
GeForce RTX 3090 Ti
No benchmark results yet

Artificial Intelligence Tests

Performance in machine learning and artificial intelligence tasks

GeekBench 6 ML

Tests throughput of AI operations in single, half, and quantized precision
GB6 ML Single Precision
RTX 4090 +61%
42962
26701
GB6 ML Half Precision
RTX 4090 +55%
59636
38573
GB6 ML Quantized
RTX 4090 +98%
31670
15995
Image Classification (SP) 15600 10246
Image Segmentation (HP) 36657 22830
Image Super Resolution (Q) 48114 23532
Face Detection (HP) 78111 48657
Pose Estimation (Q) 289639 147241
Text Classification (SP) 4074 2917
Machine Translation (HP) 6349 4427
Object Detection (SP) 21186 13323
Depth Estimation (Q) 75599 40055
Style Transfer (SP) 601428 298683
Framework ONNX ONNX
Backend DirectML DirectML
Sources: Geekbench [9], [10]

Specifications

Technical specifications of GeForce RTX 4090 and 3090 Ti

General

Vendor Nvidia Nvidia
Build Discrete Discrete
Released September 20, 2022 March 29, 2022
Launch price (MSRP) $1599 $1999
Case Desktop Desktop
Purpose Gaming Gaming
Segment High-end High-end
Architecture Ada Lovelace Ampere
GPU Codename AD102 GA102
Successor - GeForce RTX 5090 -
Recommended CPU - Intel Core i9 14900K or above - Intel Core i9 12900K or above
Desktop GPU rating (3rd and 14th place)

Graphics Processing Unit

Base Clock 2235 MHz 1560 MHz
Boost Clock 2520 MHz 1860 MHz
Shading Units 16384 10752
Texture Mapping Units (TMUs) 512 336
Render Output Units (ROPs) 176 112
Compute Units (Pipelines) 128 84
Tensor Cores 512 336
Ray-tracing Cores 128 84
L1 Cache 128KB per cluster 128KB per cluster
L2 Cache 72MB shared 6MB shared
Instructions Per Cycle 2 IPC 2 IPC

Raw Performance

Pixel Fill Rate 444 GPixel/s 208 GPixel/s
Texture Fill Rate 1290 GTexel/s 625 GTexel/s
FLOPS (FP32)
RTX 4090 +107%
82.6 TFLOPS
40 TFLOPS

Physical

Interface PCIe 4.0 x16 PCIe 4.0 x16
TGP 450 W 450 W
Manufacturing TSMC Samsung
Fabrication Process 5 nm 8 nm
Die Size 609 mm² 628 mm²
Transistor Count 76.3 billion 28 billion
Transistor Density 125.29 MTr/mm² 44.59 MTr/mm²

Memory

Memory Type GDDR6X GDDR6X
Memory Size 24 GB 24 GB
Memory Clock 2625 MHz 1313 MHz
Effective Memory Speed 21000 Mbps 21000 Mbps
Bus 384-bit 384-bit
ECC No No
Memory Bandwidth
1010 GB/s
1008 GB/s

API

DirectX 12 12
Vulkan 1.3 1.3
OpenGL 4.6 4.6
OpenCL 3.0 3.0
CUDA 8.9 8.6
Ray Tracing Yes Yes
DLSS DLSS 3 DLSS 2
DisplayPort 1.4a 1.4a

Cast your vote

Choose between two graphics cards
57 (90.5%)
6 (9.5%)
Total votes: 63

User opinions

You can share your opinion or ask a question in the comments below
🌐 Register your profile and become part of NanoReview community!