Nvidia N1X (40SM) vs N1 (20SM)
Nvidia N1X (40SM)
GB20B
Nvidia N1 (20SM)
GB20B
We performed a head-to-head comparison of the Nvidia N1X (40SM) with 40 pipelines and 5120 shaders against the N1 (20SM) that utilizes 20 pipelines and 2560 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.
Please note that the tests on the Nvidia N1X (40SM) are done on an engineering sample provided by our insiders. The data will be more accurate after the final version of this GPU is available.
Key differences
Reasons to consider the Nvidia N1X (40SM)
- 2x higher maximum theoretical performance (24.8 vs 12.4 TFLOPS)
- Manufactured using a more efficient 3 nm process technology
- Features 80 more tensor cores for effective ML and AI workloads
- Has 2x more shading units (5120 vs 2560)
Benchmarks
Graphics cards’ performance in recent benchmarking appsSpecifications
Technical specifications of Nvidia N1X (40SM) and N1 (20SM)General
| Vendor | Nvidia | Nvidia |
| Build | Integrated | Discrete |
| Released | June 1, 2026 | June 1, 2026 |
| Case | Laptop | Laptop |
| Purpose | Professional | Professional |
| Segment | Mid-range | Entry-level |
| Architecture | Blackwell (N1x) | Blackwell (N1x) |
| GPU Codename | GB20B | GB20B |
| Rival Equivalent | - Apple M5 Max GPU (32-core) | - |
| Recommended CPU | - Nvidia RTX Spark N1X (18-Core) or above | - |
Graphics Processing Unit
| Base Clock | 1665 MHz | 1665 MHz |
| Boost Clock | 2418 MHz | 2418 MHz |
| Shading Units | 5120 | 2560 |
| Texture Mapping Units (TMUs) | 160 | 160 |
| Render Output Units (ROPs) | 40 | 24 |
| Compute Units (Pipelines) | 40 | 20 |
| Tensor Cores | 160 | 80 |
| Ray-tracing Cores | 40 | 20 |
| L1 Cache | 128KB per cluster | 128KB per cluster |
| L2 Cache | 24MB shared | 50MB shared |
| Instructions Per Cycle | 2 IPC | 2 IPC |
Raw Performance
| Pixel Fill Rate | 97 GPixel/s | 58 GPixel/s |
| Texture Fill Rate | 387 GTexel/s | 387 GTexel/s |
FLOPS (FP32)
Physical
| Interface | PCIe 5.0 x16 | PCIe 5.0 x16 |
| Manufacturing | TSMC | TSMC |
| Fabrication Process | 3 nm | 5 nm |
| Die Size | 173.6 mm² | - |
Memory
| Memory Type | System Shared | System Shared |
| Bus | 256-bit | 256-bit |
| ECC | No | No |
Memory Bandwidth
API
| Vulkan | 1.4 | - |
| OpenCL | 3.0 | 3.0 |
| CUDA | 12.1 | 12.1 |
| Ray Tracing | Yes | Yes |
| DLSS | DLSS 4 | - |
Cast your vote
Total votes: 1