GeForce RTX 4070 SUPER vs 3080 Ti
We compared two discrete desktop gaming GPUs: the GeForce RTX 4070 SUPER with 56 pipelines and 7168 shaders against the 3 years and 1 month older RTX 3080 Ti that utilizes 80 pipelines and 10240 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.
Review
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score
Value for money (Beta)
Key differences
Reasons to consider the GeForce RTX 4070 SUPER
- Manufactured using a more efficient 5 nm process technology
- Supports Nvidia DLSS 3 technology
Reasons to consider the GeForce RTX 3080 Ti
- Has 81% higher memory bandwidth: 912 vs 504.2 GB/s
- Features 96 more tensor cores for effective ML and AI workloads
- Has 43% more shading units (10240 vs 7168)
Gaming Performance
Frame rate comparison across popular AAA titles at different resolutionsGames
| Forza Horizon 5 | - | 186 125 117 90 |
| The Witcher 3 | - | 204 173 125 67 |
| Counter-Strike 2 | - | 277 205 155 82 |
| Far Cry 6 | - | 153 142 123 75 |
| Hogwarts Legacy | - | 131 108 84 46 |
| CoD: Modern Warfare III | - | 169 158 121 82 |
| Ghost of Tsushima | - | 114 96 82 51 |
| Cyberpunk 2077 | - | 146 128 86 40 |
| Shadow of the Tomb Raider | - | 235 220 165 89 |
| 1080p High | - | 179 |
| 1080p Ultra | - | 151 |
| 1440p Ultra | - | 118 |
| 4K Ultra | - | 69 |
| Margin of Error | High | Low |
Benchmarks
Graphics cards’ performance in recent benchmarking apps3D Mark
Steel Nomad Lite Score
| Time Spy | 20879 | 19633 |
| Solar Bay | 98188 | 96626 |
| Port Royal | 13123 | 13265 |
| Fire Strike | 50092 | 47869 |
| Wild Life Extreme | 41625 | 42724 |
| Night Raid | 164614 | 145258 |
GeekBench 6 OpenCL
GB6 Compute Score
| Background Blur | 359.5 img/sec | 239.8 img/sec |
| Face Detection | 210 img/sec | 156.8 img/sec |
| Horizon Detection | 7.61 Gpixels/sec | 10.8 Gpixels/sec |
| Edge Detection | 10.4 Gpixels/sec | 14.8 Gpixels/sec |
| Gaussian Blur | 11.3 Gpixels/sec | 11.9 Gpixels/sec |
| Feature Matching | 1.97 Gpixels/sec | 1.53 Gpixels/sec |
| Stereo Matching | 969.3 Gpixels/sec | 889.6 Gpixels/sec |
| Particle Physics | 29061.4 FPS | 26054.6 FPS |
| API | OpenCL | OpenCL |
Cinebench 2024 GPU
Cinebench 2024 GPU
Passmark Graphics
G3D Mark Score
| G2D Mark | 1193 | 1095 |
| DirectX 11 | 271 FPS | 221 FPS |
| DirectX 12 | 110 FPS | 109 FPS |
| GPU Compute | 16925 Ops/s | 14943 Ops/s |
Blender
Blender GPU
Recent User Tests
GeForce RTX 4070 SUPER
| Date | Benchmark | Result |
|---|---|---|
| 📘 2025-12-08 (Rickey) | Cinebench 2024 | 19110 |
| 📘 2025-11-28 (Caleb) | Geekbench 6 OpenCL | 212148 |
| 📘 2025-11-28 (Caleb) | Cinebench 2024 | 20703 |
GeForce RTX 3080 Ti
| Date | Benchmark | Result |
|---|---|---|
| 📘 2025-11-01 (cancel) | Cinebench 2024 | 15466 |
Artificial Intelligence Tests
Performance in machine learning and artificial intelligence tasksGeekBench 6 ML
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
| Image Classification (SP) | 12146 | 9747 |
| Image Segmentation (HP) | 23932 | 26654 |
| Image Super Resolution (Q) | 36321 | 21475 |
| Face Detection (HP) | 52836 | 49813 |
| Pose Estimation (Q) | 179788 | 135370 |
| Text Classification (SP) | 3371 | 2796 |
| Machine Translation (HP) | 5017 | 4757 |
| Object Detection (SP) | 15164 | 12924 |
| Depth Estimation (Q) | 62500 | 36756 |
| Style Transfer (SP) | 357489 | 274032 |
| Framework | ONNX | ONNX |
| Backend | DirectML | DirectML |
Specifications
Technical specifications of GeForce RTX 4070 SUPER and 3080 TiGeneral
| Vendor | Nvidia | Nvidia |
| Build | Discrete | Discrete |
| Released | January 17, 2024 | January 3, 2021 |
| Launch price (MSRP) | $599 | $1199 |
| Case | Desktop | Desktop |
| Purpose | Gaming | Gaming |
| Segment | Mid-range | High-end |
| Architecture | Ada Lovelace | Ampere |
| GPU Codename | AD104 | GA102 |
| Rival Equivalent | - Radeon RX 7900 XT | - Radeon RX 6900 XT |
| Recommended CPU | - Intel Core i7 14700K or above | - Intel Core i9 12900K or above |
Desktop GPU rating (19th and 23rd place)
Graphics Processing Unit
| Base Clock | 1980 MHz | 1365 MHz |
| Boost Clock | 2475 MHz | 1665 MHz |
| Shading Units | 7168 | 10240 |
| Texture Mapping Units (TMUs) | 224 | 320 |
| Render Output Units (ROPs) | 80 | 112 |
| Compute Units (Pipelines) | 56 | 80 |
| Tensor Cores | 224 | 320 |
| Ray-tracing Cores | 56 | 80 |
| L1 Cache | 128KB per cluster | 128KB per cluster |
| L2 Cache | 48MB shared | 6MB shared |
| Instructions Per Cycle | 2 IPC | 2 IPC |
Raw Performance
| Pixel Fill Rate | 198 GPixel/s | 186 GPixel/s |
| Texture Fill Rate | 554 GTexel/s | 533 GTexel/s |
FLOPS (FP32)
Physical
| Interface | PCIe 4.0 x16 | PCIe 4.0 x16 |
| TGP | 220 W | 350 W |
| Manufacturing | TSMC | Samsung |
| Fabrication Process | 5 nm | 8 nm |
| Die Size | 294 mm² | 628 mm² |
| Transistor Count | 35 billion | 28 billion |
| Transistor Density | 119.05 MTr/mm² | 44.59 MTr/mm² |
Memory
| Memory Type | GDDR6X | GDDR6X |
| Memory Size | 12 GB | 12 GB |
| Memory Clock | 1313 MHz | 1188 MHz |
| Effective Memory Speed | 21000 Mbps | 19000 Mbps |
| Bus | 192-bit | 384-bit |
| ECC | No | No |
Memory Bandwidth
API
| DirectX | 12 | 12 |
| Vulkan | 1.3 | 1.3 |
| OpenGL | 4.6 | 4.6 |
| OpenCL | 3.0 | 3.0 |
| CUDA | 8.9 | 8.6 |
| Ray Tracing | Yes | Yes |
| DLSS | DLSS 3 | DLSS 2 |
| DisplayPort | 1.4a | 1.4a |
Cast your vote
Total votes: 465