Nvidia N1 (16SM) vs Apple M4 GPU (10-Core)

We performed a head-to-head comparison of the Nvidia N1 (16SM) with 16 pipelines and 2048 shaders against the 2 years and 2 months older Apple M4 GPU (10-Core) that utilizes 160 pipelines and 1280 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.

Review

General comparison of performance in games, applications, power efficiency, and other metrics
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score
Scores marked with a red asterisk (**) represent initial estimates from early testing.

Key differences

Key distinctions and advantages of M4 GPU (10-Core) over N1 (16SM)
Reasons to consider the Nvidia N1 (16SM)
  • 2.6x higher maximum theoretical performance (9.9 vs 3.8 TFLOPS)
  • Has 2.3x higher memory bandwidth: 273.2 vs 120 GB/s
  • Has 60% more shading units (2048 vs 1280)
Reasons to consider the Apple M4 GPU (10-Core)
  • Manufactured using a more efficient 3 nm process technology

Benchmarks

Graphics cards’ performance in recent benchmarking apps

3D Mark

Multiplatform graphics benchmark suite that directly correlates with performance in modern games
Steel Nomad Lite Score
Solar Bay - 16554
Wild Life Extreme - 9610
Sources: 3DMark [1]

GeekBench 6 GPU Compute

GPU test for computational tasks (image processing, photography, computer vision, and ML)
GB6 Compute Score
Background Blur - 72.2 img/sec
Face Detection - 48.3 img/sec
Horizon Detection - 1.4 Gpixels/sec
Edge Detection - 1.94 Gpixels/sec
Gaussian Blur - 1.5 Gpixels/sec
Feature Matching - 0.53 Gpixels/sec
Stereo Matching - 123.3 Gpixels/sec
Particle Physics - 4938.4 FPS
API - OpenCL
Sources: Geekbench [3]

Cinebench 2024 GPU

Hardware benchmark using Maxon's Cinema 4D rendering engine
Cinebench 2024 GPU

Blender

Rendering performance test for 3D modeling
Blender GPU
Sources: Blender [9] – 551 samples

Artificial Intelligence Tests

Performance in machine learning and artificial intelligence tasks

GeekBench 6 ML

Tests throughput of AI operations in single, half, and quantized precision
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
Image Classification (SP) - 4874
Image Segmentation (HP) - 8168
Image Super Resolution (Q) - 11983
Face Detection (HP) - 17036
Pose Estimation (Q) - 31171
Text Classification (SP) - 2964
Machine Translation (HP) - 3410
Object Detection (SP) - 5112
Depth Estimation (Q) - 22700
Style Transfer (SP) - 60872
Framework - Core ML
Backend - GPU
Sources: Geekbench [9]

Specifications

Technical specifications of Nvidia N1 (16SM) and Apple M4 GPU (10-Core)

General

Vendor Nvidia Apple
Build Discrete Integrated
Released June 1, 2026 May 7, 2024
Case Laptop Laptop
Purpose Professional Professional
Segment Entry-level Mid-range
Architecture Blackwell (N1x) Apple M GPU
GPU Codename GB20B Custom
Rival Equivalent - - Adreno X1-85
Successor - - Apple M5 GPU (10-Core)
Recommended CPU - - Apple M4 (10-Core) or above
Used in CPUs - - Apple M4 (10-Core)
Laptop GPU ranking (18th and 73rd place)

Graphics Processing Unit

Base Clock 1665 MHz 500 MHz
Boost Clock 2418 MHz 1470 MHz
Shading Units 2048 1280
Texture Mapping Units (TMUs) 128 80
Render Output Units (ROPs) 24 40
Compute Units (Pipelines) 16 160
Tensor Cores 64 -
Ray-tracing Cores 16 -
L1 Cache 128KB per cluster -
L2 Cache 50MB shared -
Instructions Per Cycle 2 IPC 2 IPC

Raw Performance

Pixel Fill Rate 58 GPixel/s 59 GPixel/s
Texture Fill Rate 310 GTexel/s 118 GTexel/s
FLOPS (FP32)
N1 (16SM) +161%
9.9 TFLOPS
3.8 TFLOPS

Physical

Interface PCIe 5.0 x16 Custom
TGP - 18 W
Manufacturing TSMC TSMC
Fabrication Process 5 nm 3 nm
Transistor Count - 28 billion
Max. Temperature - 100°C

Memory

Memory Type System Shared System Shared
Memory Clock - 7500 MHz
Effective Memory Speed - 15000 Mbps
Bus 256-bit 128-bit
ECC No No
Memory Bandwidth
N1 (16SM) +128%
273.2 GB/s
120 GB/s

API

OpenCL 3.0 -
CUDA 12.1 -
Ray Tracing Yes Yes
DLSS - No

Cast your vote

Choose between two graphics cards
0 (0%)
0 (0%)
Total votes: < 1

User opinions

You can share your opinion or ask a question in the comments below
🌐 Register your profile and become part of NanoReview community!