Nvidia N1X (40SM) vs N1 (16SM)
Nvidia N1X (40SM)
GB20B
Nvidia N1 (16SM)
GB20B
We performed a head-to-head comparison of the Nvidia N1X (40SM) with 40 pipelines and 5120 shaders against the N1 (16SM) that utilizes 16 pipelines and 2048 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.
Please note that the tests on the Nvidia N1X (40SM) are done on an engineering sample provided by our insiders. The data will be more accurate after the final version of this GPU is available.
Key differences
Reasons to consider the Nvidia N1X (40SM)
- 2.5x higher maximum theoretical performance (24.8 vs 9.9 TFLOPS)
- Manufactured using a more efficient 3 nm process technology
- Features 96 more tensor cores for effective ML and AI workloads
- Has 2.5x more shading units (5120 vs 2048)
Benchmarks
Graphics cards’ performance in recent benchmarking appsSpecifications
Technical specifications of Nvidia N1X (40SM) and N1 (16SM)General
| Vendor | Nvidia | Nvidia |
| Build | Integrated | Discrete |
| Released | June 1, 2026 | June 1, 2026 |
| Case | Laptop | Laptop |
| Purpose | Professional | Professional |
| Segment | Mid-range | Entry-level |
| Architecture | Blackwell (N1x) | Blackwell (N1x) |
| GPU Codename | GB20B | GB20B |
| Rival Equivalent | - Apple M5 Max GPU (32-core) | - |
| Recommended CPU | - Nvidia RTX Spark N1X (18-Core) or above | - |
Laptop GPU ranking (#18th place)
Graphics Processing Unit
| Base Clock | 1665 MHz | 1665 MHz |
| Boost Clock | 2418 MHz | 2418 MHz |
| Shading Units | 5120 | 2048 |
| Texture Mapping Units (TMUs) | 160 | 128 |
| Render Output Units (ROPs) | 40 | 24 |
| Compute Units (Pipelines) | 40 | 16 |
| Tensor Cores | 160 | 64 |
| Ray-tracing Cores | 40 | 16 |
| L1 Cache | 128KB per cluster | 128KB per cluster |
| L2 Cache | 24MB shared | 50MB shared |
| Instructions Per Cycle | 2 IPC | 2 IPC |
Raw Performance
| Pixel Fill Rate | 97 GPixel/s | 58 GPixel/s |
| Texture Fill Rate | 387 GTexel/s | 310 GTexel/s |
FLOPS (FP32)
Physical
| Interface | PCIe 5.0 x16 | PCIe 5.0 x16 |
| Manufacturing | TSMC | TSMC |
| Fabrication Process | 3 nm | 5 nm |
| Die Size | 173.6 mm² | - |
Memory
| Memory Type | System Shared | System Shared |
| Bus | 256-bit | 256-bit |
| ECC | No | No |
Memory Bandwidth
API
| Vulkan | 1.4 | - |
| OpenCL | 3.0 | 3.0 |
| CUDA | 12.1 | 12.1 |
| Ray Tracing | Yes | Yes |
| DLSS | DLSS 4 | - |
Cast your vote
Total votes: 5