GeForce RTX 5070 vs 4060 Ti 8GB
Nvidia GeForce RTX 5070
GB205-300
We compared two discrete desktop gaming GPUs: the GeForce RTX 5070 12 GB with 48 pipelines and 6144 shaders against the 1 year and 8 months older RTX 4060 Ti 8GB 8 GB that utilizes 34 pipelines and 4352 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.
Review
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score
Value for money (Beta)
Key differences
Reasons to consider the GeForce RTX 5070
- Performs significantly better (up to 71%) in 3DMark Steel Nomad Lite
- Shows 49% higher average frame rate in modern games at QHD resolution – 128 vs 86 FPS
- 40% higher maximum theoretical performance (30.9 vs 22.1 TFLOPS)
- Manufactured using a more efficient 4 nm process technology
Gaming Performance
Frame rate comparison across popular AAA titles at different resolutionsGames
| Forza Horizon 5 | 197 140 130 97 |
150 97 84 61 |
| The Witcher 3 | 216 183 131 71 |
167 137 98 52 |
| Counter-Strike 2 | 293 210 160 81 |
239 187 136 76 |
| Far Cry 6 | 164 145 128 82 |
129 119 95 54 |
| Hogwarts Legacy | 140 117 92 51 |
97 80 59 34 |
| CoD: Modern Warfare III | 178 171 136 89 |
121 112 80 53 |
| Ghost of Tsushima | 123 107 89 58 |
86 74 56 33 |
| Cyberpunk 2077 | 164 144 99 43 |
102 91 57 26 |
| Shadow of the Tomb Raider | 253 238 189 98 |
164 146 109 58 |
| 1080p High | 192 | 139 |
| 1080p Ultra | 162 | 116 |
| 1440p Ultra | 128 | 86 |
| 4K Ultra | 74 | 50 |
| Margin of Error | Low | Low |
Benchmarks
Graphics cards’ performance in recent benchmarking apps3D Mark
Steel Nomad Lite Score
| Time Spy | 22708 | 13479 |
| Solar Bay | 107004 | 62428 |
| Port Royal | 14498 | 8103 |
| Fire Strike | 58324 | 34401 |
| Wild Life Extreme | 43879 | 25599 |
| Night Raid | 172064 | 136255 |
GeekBench 6 OpenCL
GB6 Compute Score
| Background Blur | 254.4 img/sec | 198.3 img/sec |
| Face Detection | 180.6 img/sec | 126.2 img/sec |
| Horizon Detection | 7.91 Gpixels/sec | 4.35 Gpixels/sec |
| Edge Detection | 10.2 Gpixels/sec | 5.69 Gpixels/sec |
| Gaussian Blur | 10.7 Gpixels/sec | 6.72 Gpixels/sec |
| Feature Matching | 1.84 Gpixels/sec | 1.53 Gpixels/sec |
| Stereo Matching | 975.7 Gpixels/sec | 656.3 Gpixels/sec |
| Particle Physics | 27816.5 FPS | 19149.2 FPS |
| API | OpenCL | OpenCL |
Cinebench 2024 GPU
Cinebench 2024 GPU
Passmark Graphics
G3D Mark Score
| G2D Mark | 1289 | 1094 |
| DirectX 11 | 279 FPS | 202 FPS |
| DirectX 12 | 106 FPS | 85 FPS |
| GPU Compute | 14323 Ops/s | 12003 Ops/s |
Artificial Intelligence Tests
Performance in machine learning and artificial intelligence tasksGeekBench 6 ML
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
| Image Classification (SP) | 12270 | 9359 |
| Image Segmentation (HP) | 30237 | 21463 |
| Image Super Resolution (Q) | 35585 | 26023 |
| Face Detection (HP) | 62662 | 42533 |
| Pose Estimation (Q) | 145190 | 99698 |
| Text Classification (SP) | 4198 | 3181 |
| Machine Translation (HP) | 6233 | 4956 |
| Object Detection (SP) | 16709 | 11009 |
| Depth Estimation (Q) | 58300 | 43518 |
| Style Transfer (SP) | 246598 | 203667 |
| Framework | ONNX | ONNX |
| Backend | DirectML | DirectML |
Specifications
Technical specifications of GeForce RTX 5070 and 4060 Ti 8GBGeneral
| Vendor | Nvidia | Nvidia |
| Build | Discrete | Discrete |
| Released | January 7, 2025 | May 18, 2023 |
| Launch price (MSRP) | $549 | $399 |
| Case | Desktop | Desktop |
| Purpose | Gaming | Gaming |
| Segment | Mid-range | Mid-range |
| Architecture | Blackwell 2.0 | Ada Lovelace |
| GPU Codename | GB205-300 | AD106 |
| Rival Equivalent | - Radeon RX 9070 | - Radeon RX 7700 XT |
| Successor | - | - GeForce RTX 5060 Ti (8GB) |
| Recommended CPU | - Intel Core Ultra 5 245K or above | - Intel Core i5 14600K or above |
Desktop GPU rating (21st and 43rd place)
Graphics Processing Unit
| Base Clock | 2160 MHz | 2310 MHz |
| Boost Clock | 2512 MHz | 2535 MHz |
| Shading Units | 6144 | 4352 |
| Texture Mapping Units (TMUs) | 192 | 136 |
| Render Output Units (ROPs) | 80 | 48 |
| Compute Units (Pipelines) | 48 | 34 |
| Tensor Cores | 192 | 136 |
| Ray-tracing Cores | 48 | 34 |
| L1 Cache | 128KB per cluster | 128KB per cluster |
| L2 Cache | 48MB shared | 32MB shared |
| Instructions Per Cycle | 2 IPC | 2 IPC |
Raw Performance
| Pixel Fill Rate | 201 GPixel/s | 122 GPixel/s |
| Texture Fill Rate | 482 GTexel/s | 345 GTexel/s |
FLOPS (FP32)
Physical
| Interface | PCIe 5.0 x16 | PCIe 4.0 x8 |
| TGP | 250 W | 160 W |
| Manufacturing | TSMC | TSMC |
| Fabrication Process | 4 nm | 5 nm |
| Die Size | 263 mm² | 188 mm² |
| Transistor Count | 31.1 billion | 22 billion |
| Transistor Density | 118.25 MTr/mm² | 117.02 MTr/mm² |
Memory
| Memory Type | GDDR7 | GDDR6 |
| Memory Size | 12 GB | 8 GB |
| Memory Clock | 1750 MHz | 2250 MHz |
| Effective Memory Speed | 28000 Mbps | 18000 Mbps |
| Bus | 192-bit | 128-bit |
| ECC | No | No |
Memory Bandwidth
API
| DirectX | 12.2 | 12 |
| Vulkan | 1.4 | 1.3 |
| OpenGL | 4.6 | 4.6 |
| OpenCL | 3.0 | 3.0 |
| CUDA | 12.0 | 8.9 |
| Ray Tracing | Yes | Yes |
| DLSS | DLSS 4 | DLSS 3 |
| DisplayPort | 2.1b | 1.4a |
Cast your vote
Total votes: 41