Apple M1 Pro GPU (16-core) vs GPU (8-core)

We compared two integrated laptop professional GPUs: the Apple M1 Pro GPU (16-core) with 256 pipelines and 2048 shaders against the 11 months older GPU (8-core) that utilizes 128 pipelines and 1024 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.

Review

General comparison of performance in games, applications, power efficiency, and other metrics
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
AI/ML
Capabilities for machine learning and AI-related tasks
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score

Key differences

Key distinctions and advantages of M1 GPU (8-core) over M1 Pro GPU (16-core)
Reasons to consider the Apple M1 Pro GPU (16-core)
  • Performs significantly better (up to 2x) in 3DMark Steel Nomad Lite
  • 2x higher maximum theoretical performance (5.3 vs 2.6 TFLOPS)
  • Has 3x higher memory bandwidth: 204.8 vs 68.2 GB/s
  • Achieves 2x more points in the GeekBench 6 Compute test (42K vs 20K)
  • Has 2x more shading units (2048 vs 1024)

Benchmarks

Graphics cardsโ€™ performance in recent benchmarking apps

3D Mark

Multiplatform graphics benchmark suite that directly correlates with performance in modern games
Steel Nomad Lite Score
Solar Bay 12485 6317
Wild Life Extreme 9873 4966
Sources: 3DMark [1], [2]

GeekBench 6 OpenCL

GPU test for computational tasks (image processing, photography, computer vision, and ML)
GB6 Compute Score
Background Blur 74.5 img/sec 38.3 img/sec
Face Detection 47.3 img/sec 23 img/sec
Horizon Detection 1.73 Gpixels/sec 0.76 Gpixels/sec
Edge Detection 3.32 Gpixels/sec 1.33 Gpixels/sec
Gaussian Blur 1.75 Gpixels/sec 0.87 Gpixels/sec
Feature Matching 0.47 Gpixels/sec 0.27 Gpixels/sec
Stereo Matching 131.2 Gpixels/sec 69.9 Gpixels/sec
Particle Physics 5143 FPS 2766.4 FPS
API OpenCL OpenCL
Sources: Geekbench [3], [4]

Cinebench 2024 GPU

Hardware benchmark using Maxon's Cinema 4D rendering engine
Cinebench 2024 GPU

Blender

Rendering performance test for 3D modeling
Blender GPU
Sources: Blender [9], [10] โ€“ 502 & 307 samples

Recent User Tests

The latest benchmark tests that have been submitted by users
Apple M1 Pro GPU (16-core)
DateBenchmarkResult
๐Ÿ“˜ 2026-03-27 (bbffx)Geekbench 6 OpenCL43467
๐Ÿ“˜ 2026-03-27 (bbffx)Cinebench 20242423
๐Ÿ“˜ 2026-03-27 (bbffx)Steel Nomad Light4048
๐Ÿ“˜ 2025-02-21 (serhii)Geekbench 6 OpenCL40932
Apple M1 GPU (8-core)
No benchmark results yet

Artificial Intelligence Tests

Performance in machine learning and artificial intelligence tasks

GeekBench 6 ML

Tests throughput of AI operations in single, half, and quantized precision
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
Image Classification (SP) 4112 1400
Image Segmentation (HP) 7488 2280
Image Super Resolution (Q) 9882 3402
Face Detection (HP) 14702 5638
Pose Estimation (Q) 33560 12027
Text Classification (SP) 2289 2324
Machine Translation (HP) 1629 4296
Object Detection (SP) 3633 1530
Depth Estimation (Q) 16986 7675
Style Transfer (SP) 53586 17293
Framework Core ML Core ML
Backend GPU CPU
Sources: Geekbench [9], [10]

Specifications

Technical specifications of Apple M1 Pro GPU (16-core) and GPU (8-core)

General

Vendor Apple Apple
Build Integrated Integrated
Released October 18, 2021 November 20, 2020
Case Laptop Laptop
Purpose Professional Professional
Segment Mid-range Mid-range
Architecture Apple M GPU Apple M GPU
Rival Equivalent - GeForce RTX 3050 Laptop - GeForce GTX 1050 Ti Mobile
Successor - Apple M5 Pro GPU (20-core) - Apple M5 GPU (10-Core)
Recommended CPU - Apple M1 Pro or above - Apple M1 or above
Used in CPUs - Apple M1 Pro - Apple M1
Laptop GPU ranking (78th and 104th place)

Graphics Processing Unit

Base Clock 450 MHz 450 MHz
Boost Clock 1296 MHz 1278 MHz
Shading Units 2048 1024
Texture Mapping Units (TMUs) 128 64
Render Output Units (ROPs) 64 32
Compute Units (Pipelines) 256 128
Ray-tracing Cores No No
Instructions Per Cycle 2 IPC 2 IPC

Raw Performance

Pixel Fill Rate 83 GPixel/s 41 GPixel/s
Texture Fill Rate 166 GTexel/s 82 GTexel/s
FLOPS (FP32)
5.3 TFLOPS
2.6 TFLOPS

Physical

Interface Custom Custom
TGP 30 W 15 W
Manufacturing TSMC Intel
Fabrication Process 5 nm 5 nm
Transistor Count 23.2 billion 11.6 billion
Max. Temperature 94ยฐC 94ยฐC

Memory

Memory Type System Shared System Shared
Memory Clock 6400 MHz 4266 MHz
Effective Memory Speed 12800 Mbps 8532 Mbps
Bus 256-bit 128-bit
ECC No No
Memory Bandwidth
204.8 GB/s
68.2 GB/s

API

Ray Tracing No No
DLSS No No

Cast your vote

Choose between two graphics cards
6 (100%)
0 (0%)
Total votes: 6

User opinions

You can share your opinion or ask a question in the comments below
๐ŸŒ Register your profile and become part of NanoReview community!