GeForce RTX 4080 SUPER vs 4070 Ti SUPER
We compared two discrete desktop gaming GPUs: the GeForce RTX 4080 SUPER with 80 pipelines and 10240 shaders against the RTX 4070 Ti SUPER that utilizes 66 pipelines and 8448 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.
Review
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score
Value for money (Beta)
Key differences
Reasons to consider the GeForce RTX 4080 SUPER
- Performs better (up to 19%) in 3DMark Steel Nomad Lite
- Shows 11% higher average frame rate in modern games at QHD resolution – 149 vs 134 FPS
- 18% higher maximum theoretical performance (52.2 vs 44.1 TFLOPS)
Gaming Performance
Frame rate comparison across popular AAA titles at different resolutionsGames
| Forza Horizon 5 | 231 166 153 123 |
210 148 140 107 |
| The Witcher 3 | 242 215 158 83 |
221 192 140 72 |
| Counter-Strike 2 | 302 227 167 85 |
294 219 164 83 |
| Far Cry 6 | 182 162 146 92 |
168 156 138 83 |
| Hogwarts Legacy | 166 136 110 60 |
146 122 94 54 |
| CoD: Modern Warfare III | 216 208 168 115 |
195 189 144 99 |
| Ghost of Tsushima | 146 125 111 71 |
134 113 95 62 |
| Cyberpunk 2077 | 199 174 123 59 |
169 153 105 49 |
| Shadow of the Tomb Raider | 288 266 201 112 |
270 250 190 106 |
| 1080p High | 219 | 201 |
| 1080p Ultra | 187 | 171 |
| 1440p Ultra | 149 | 134 |
| 4K Ultra | 89 | 79 |
| Margin of Error | Low | Low |
Benchmarks
Graphics cards’ performance in recent benchmarking apps3D Mark
Steel Nomad Lite Score
| Time Spy | 28271 | 24220 |
| Solar Bay | 139050 | 116588 |
| Port Royal | 18353 | 15806 |
| Fire Strike | 63706 | 56343 |
| Wild Life Extreme | 59996 | 49911 |
| Night Raid | 180569 | 176672 |
GeekBench 6 OpenCL
GB6 Compute Score
| Background Blur | 279.5 img/sec | 331.7 img/sec |
| Face Detection | 220.1 img/sec | 211.6 img/sec |
| Horizon Detection | 11 Gpixels/sec | 9.63 Gpixels/sec |
| Edge Detection | 15.2 Gpixels/sec | 13.7 Gpixels/sec |
| Gaussian Blur | 16.2 Gpixels/sec | 13.9 Gpixels/sec |
| Feature Matching | 2.08 Gpixels/sec | 2.15 Gpixels/sec |
| Stereo Matching | 1370 Gpixels/sec | 1210 Gpixels/sec |
| Particle Physics | 37199.7 FPS | 33566 FPS |
| API | OpenCL | OpenCL |
Cinebench 2024 GPU
Cinebench 2024 GPU
Passmark Graphics
G3D Mark Score
| G2D Mark | 1278 | 1246 |
| DirectX 11 | 299 FPS | 277 FPS |
| DirectX 12 | 135 FPS | 120 FPS |
| GPU Compute | 19585 Ops/s | 18196 Ops/s |
Blender
Blender GPU
Artificial Intelligence Tests
Performance in machine learning and artificial intelligence tasksGeekBench 6 ML
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
| Image Classification (SP) | 12453 | 12970 |
| Image Segmentation (HP) | 34467 | 26364 |
| Image Super Resolution (Q) | 43534 | 37217 |
| Face Detection (HP) | 70300 | 58720 |
| Pose Estimation (Q) | 242331 | 177604 |
| Text Classification (SP) | 3767 | 3815 |
| Machine Translation (HP) | 6614 | 5829 |
| Object Detection (SP) | 18301 | 15675 |
| Depth Estimation (Q) | 66477 | 62447 |
| Style Transfer (SP) | 417906 | 357918 |
| Framework | ONNX | ONNX |
| Backend | DirectML | DirectML |
Specifications
Technical specifications of GeForce RTX 4080 SUPER and 4070 Ti SUPERGeneral
| Vendor | Nvidia | Nvidia |
| Build | Discrete | Discrete |
| Released | January 31, 2024 | January 24, 2024 |
| Launch price (MSRP) | $999 | $799 |
| Case | Desktop | Desktop |
| Purpose | Gaming | Gaming |
| Segment | High-end | High-end |
| Architecture | Ada Lovelace | Ada Lovelace |
| GPU Codename | AD103 | AD103 |
| Rival Equivalent | - Radeon RX 7900 XTX | - Radeon RX 7900 XTX |
| Recommended CPU | - Intel Core i7 14700K or above | - Intel Core i7 14700K or above |
Desktop GPU rating (7th and 13th place)
Graphics Processing Unit
| Base Clock | 2295 MHz | 2340 MHz |
| Boost Clock | 2550 MHz | 2610 MHz |
| Shading Units | 10240 | 8448 |
| Texture Mapping Units (TMUs) | 320 | 264 |
| Render Output Units (ROPs) | 112 | 96 |
| Compute Units (Pipelines) | 80 | 66 |
| Tensor Cores | 320 | 264 |
| Ray-tracing Cores | 80 | 66 |
| L1 Cache | 128KB per cluster | 128KB per cluster |
| L2 Cache | 64MB shared | 48MB shared |
| Instructions Per Cycle | 2 IPC | 2 IPC |
Raw Performance
| Pixel Fill Rate | 286 GPixel/s | 251 GPixel/s |
| Texture Fill Rate | 816 GTexel/s | 689 GTexel/s |
FLOPS (FP32)
Physical
| Interface | PCIe 4.0 x16 | PCIe 4.0 x16 |
| TGP | 320 W | 285 W |
| Manufacturing | TSMC | TSMC |
| Fabrication Process | 5 nm | 5 nm |
| Die Size | 379 mm² | 379 mm² |
| Transistor Count | 45 billion | 45 billion |
| Transistor Density | 118.73 MTr/mm² | 118.73 MTr/mm² |
Memory
| Memory Type | GDDR6X | GDDR6X |
| Memory Size | 16 GB | 16 GB |
| Memory Clock | 1438 MHz | 1313 MHz |
| Effective Memory Speed | 23000 Mbps | 21000 Mbps |
| Bus | 256-bit | 256-bit |
| ECC | No | No |
Memory Bandwidth
API
| DirectX | 12 | 12 |
| Vulkan | 1.3 | 1.3 |
| OpenGL | 4.6 | 4.6 |
| OpenCL | 3.0 | 3.0 |
| CUDA | 8.9 | 8.9 |
| Ray Tracing | Yes | Yes |
| DLSS | DLSS 3 | DLSS 3 |
| DisplayPort | 1.4a | 1.4a |
Cast your vote
Total votes: 306