GeForce RTX 4070 SUPER vs 4060 Ti 16GB
We compared two discrete desktop gaming GPUs: the GeForce RTX 4070 SUPER 12 GB with 56 pipelines and 7168 shaders against the 6 months older RTX 4060 Ti 16GB 16 GB that utilizes 34 pipelines and 4352 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.
Review
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score
Scores marked with a red asterisk (**) represent initial estimates from early testing.
Value for money (Beta)
Key differences
Reasons to consider the GeForce RTX 4070 SUPER
- Performs significantly better (up to 59%) in 3DMark Steel Nomad Lite
- 61% higher maximum theoretical performance (35.5 vs 22.1 TFLOPS)
- Has 75% higher memory bandwidth: 504.2 vs 288 GB/s
- Features 88 more tensor cores for effective ML and AI workloads
- Achieves 52% more points in the GeekBench 6 Compute test (206K vs 135K)
- Has 65% more shading units (7168 vs 4352)
Reasons to consider the GeForce RTX 4060 Ti 16GB
- Includes 4 GB more video memory
Gaming Performance
Frame rate comparison across popular AAA titles at different resolutionsGames
| Forza Horizon 5 | - | 144 94 83 62 |
| The Witcher 3 | - | 172 141 99 53 |
| Counter-Strike 2 | - | 245 190 132 76 |
| Far Cry 6 | - | 129 117 92 56 |
| Hogwarts Legacy | - | 95 81 62 32 |
| CoD: Modern Warfare III | - | 118 112 81 52 |
| Ghost of Tsushima | - | 88 75 57 31 |
| Cyberpunk 2077 | - | 105 92 56 26 |
| Shadow of the Tomb Raider | - | 161 147 111 59 |
| 1080p High | - | 140 |
| 1080p Ultra | - | 117 |
| 1440p Ultra | - | 86 |
| 4K Ultra | - | 50 |
| Margin of Error | High | Low |
Benchmarks
Graphics cards’ performance in recent benchmarking apps3D Mark
Steel Nomad Lite Score
| Time Spy | 20879 | 13456 |
| Solar Bay | 98188 | 62446 |
| Port Royal | 13123 | 8119 |
| Fire Strike | 50092 | 34258 |
| Wild Life Extreme | 41625 | 25587 |
| Night Raid | 164614 | 135657 |
GeekBench 6 OpenCL
GB6 Compute Score
| Background Blur | 359.5 img/sec | 232.6 img/sec |
| Face Detection | 210 img/sec | 138.4 img/sec |
| Horizon Detection | 7.61 Gpixels/sec | 4.68 Gpixels/sec |
| Edge Detection | 10.4 Gpixels/sec | 6.06 Gpixels/sec |
| Gaussian Blur | 11.3 Gpixels/sec | 6.99 Gpixels/sec |
| Feature Matching | 1.97 Gpixels/sec | 1.6 Gpixels/sec |
| Stereo Matching | 969.3 Gpixels/sec | 652.6 Gpixels/sec |
| Particle Physics | 29061.4 FPS | 19534.7 FPS |
| API | OpenCL | OpenCL |
Cinebench 2024 GPU
Cinebench 2024 GPU
Passmark Graphics
G3D Mark Score
| G2D Mark | 1193 | 1078 |
| DirectX 11 | 271 FPS | 199 FPS |
| DirectX 12 | 110 FPS | 89 FPS |
| GPU Compute | 16925 Ops/s | 11953 Ops/s |
Blender
Blender GPU
Recent User Tests
GeForce RTX 4070 SUPER
| Date | Benchmark | Result |
|---|---|---|
| 📘 2025-12-08 (Rickey) | Cinebench 2024 | 19110 |
| 📘 2025-11-28 (Caleb) | Geekbench 6 OpenCL | 212148 |
| 📘 2025-11-28 (Caleb) | Cinebench 2024 | 20703 |
GeForce RTX 4060 Ti 16GB
No benchmark results yet
Artificial Intelligence Tests
Performance in machine learning and artificial intelligence tasksGeekBench 6 ML
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
| Image Classification (SP) | 12146 | - |
| Image Segmentation (HP) | 23932 | - |
| Image Super Resolution (Q) | 36321 | - |
| Face Detection (HP) | 52836 | - |
| Pose Estimation (Q) | 179788 | - |
| Text Classification (SP) | 3371 | - |
| Machine Translation (HP) | 5017 | - |
| Object Detection (SP) | 15164 | - |
| Depth Estimation (Q) | 62500 | - |
| Style Transfer (SP) | 357489 | - |
| Framework | ONNX | - |
| Backend | DirectML | - |
Specifications
Technical specifications of GeForce RTX 4070 SUPER and 4060 Ti 16GBGeneral
| Vendor | Nvidia | Nvidia |
| Build | Discrete | Discrete |
| Released | January 17, 2024 | July 18, 2023 |
| Launch price (MSRP) | $599 | $499 |
| Case | Desktop | Desktop |
| Purpose | Gaming | Gaming |
| Segment | Mid-range | Mid-range |
| Architecture | Ada Lovelace | Ada Lovelace |
| GPU Codename | AD104 | AD106 |
| Rival Equivalent | - Radeon RX 7900 XT | - |
| Successor | - | - GeForce RTX 5060 Ti (16GB) |
| Recommended CPU | - Intel Core i7 14700K or above | - Intel Core i5 14600K or above |
Desktop GPU rating (19th and 40th place)
Graphics Processing Unit
| Base Clock | 1980 MHz | 2310 MHz |
| Boost Clock | 2475 MHz | 2535 MHz |
| Shading Units | 7168 | 4352 |
| Texture Mapping Units (TMUs) | 224 | 136 |
| Render Output Units (ROPs) | 80 | 48 |
| Compute Units (Pipelines) | 56 | 34 |
| Tensor Cores | 224 | 136 |
| Ray-tracing Cores | 56 | 34 |
| L1 Cache | 128KB per cluster | 128KB per cluster |
| L2 Cache | 48MB shared | 32MB shared |
| Instructions Per Cycle | 2 IPC | 2 IPC |
Raw Performance
| Pixel Fill Rate | 198 GPixel/s | 122 GPixel/s |
| Texture Fill Rate | 554 GTexel/s | 345 GTexel/s |
FLOPS (FP32)
Physical
| Interface | PCIe 4.0 x16 | PCIe 4.0 x8 |
| TGP | 220 W | 165 W |
| Manufacturing | TSMC | TSMC |
| Fabrication Process | 5 nm | 5 nm |
| Die Size | 294 mm² | 188 mm² |
| Transistor Count | 35 billion | 22 billion |
| Transistor Density | 119.05 MTr/mm² | 117.02 MTr/mm² |
Memory
| Memory Type | GDDR6X | GDDR6 |
| Memory Size | 12 GB | 16 GB |
| Memory Clock | 1313 MHz | 2250 MHz |
| Effective Memory Speed | 21000 Mbps | 18000 Mbps |
| Bus | 192-bit | 128-bit |
| ECC | No | No |
Memory Bandwidth
API
| DirectX | 12 | 12 |
| Vulkan | 1.3 | 1.3 |
| OpenGL | 4.6 | 4.6 |
| OpenCL | 3.0 | 3.0 |
| CUDA | 8.9 | 8.9 |
| Ray Tracing | Yes | Yes |
| DLSS | DLSS 3 | DLSS 3 |
| DisplayPort | 1.4a | 1.4a |
Cast your vote
Total votes: 83