AMD Radeon 8040S vs Apple M5 Max GPU (32-core)

We compared two integrated laptop graphics cards: the Radeon 8040S with 16 pipelines and 1024 shaders against the 1 year and 2 months newer Apple M5 Max GPU (32-core) that utilizes 512 pipelines and 4096 shaders. Here you will find complete details about specs, efficiency, performance tests, and more.

Review

General comparison of performance in games, applications, power efficiency, and other metrics
Gaming
Performance in DirectX, OpenCL, and Vulkan games
Workstation
Perf. in 3D modeling, video editing and rendering apps
Energy Efficiency
Power consumption efficiency in different scenarios
NanoReview Final Score
Overall video card score

Key differences

Key distinctions and advantages of M5 Max GPU (32-core) over Radeon 8040S
Reasons to consider the Apple M5 Max GPU (32-core)
  • Performs significantly better (up to 3.4x) in 3DMark Steel Nomad Lite
  • 2.3x higher maximum theoretical performance (13.3 vs 5.7 TFLOPS)
  • Manufactured using a more efficient 3 nm process technology
  • Achieves 3.5x more points in the GeekBench 6 Compute test (124K vs 35K)
  • Has 4x more shading units (4096 vs 1024)

Gaming Performance

Frame rate comparison across popular AAA titles at different resolutions

Games

FPS Table
Forza Horizon 5
76
40
31
17
-
The Witcher 3
53
44
26
14
-
Counter-Strike 2
84
60
39
15
-
Far Cry 6
61
49
35
16
-
Hogwarts Legacy
33
26
20
10
-
CoD: Modern Warfare III
49
41
31
15
-
Ghost of Tsushima
33
25
18
10
-
Cyberpunk 2077
32
29
17
-
-
Shadow of the Tomb Raider
65
55
32
17
-
Average FPS by Resolution
1080p High 54 -
1080p Ultra 41 -
1440p Ultra 28 -
4K Ultra 13 -
Margin of Error High High

Benchmarks

Graphics cards’ performance in recent benchmarking apps

3D Mark

Multiplatform graphics benchmark suite that directly correlates with performance in modern games
Steel Nomad Lite Score
Solar Bay - 69788
Wild Life Extreme - 35786
Sources: 3DMark [1], [2]

GeekBench 6 OpenCL

GPU test for computational tasks (image processing, photography, computer vision, and ML)
GB6 Compute Score
35976
124299
Background Blur 106.8 img/sec 234.8 img/sec
Face Detection 43.8 img/sec 158.1 img/sec
Horizon Detection 1.12 Gpixels/sec 4.48 Gpixels/sec
Edge Detection 1.17 Gpixels/sec 6.62 Gpixels/sec
Gaussian Blur 1.07 Gpixels/sec 7.19 Gpixels/sec
Feature Matching 0.5 Gpixels/sec 1.17 Gpixels/sec
Stereo Matching 163 Gpixels/sec 384.9 Gpixels/sec
Particle Physics 5820.4 FPS 17866.8 FPS
API OpenCL OpenCL
Sources: Geekbench [3], [4]

Passmark Graphics

Videocard test that focuses on compute shaders, multi-texturing, tessellation, and other features
G3D Mark Score
G2D Mark 1052 -
DirectX 11 80 FPS -
DirectX 12 47 FPS -
GPU Compute 5138 Ops/s -
Sources: PassMark [5] – 6 samples

Blender

Rendering performance test for 3D modeling
Blender GPU
376.14
6401.5
Sources: Blender [9], [10] – 1 & 37 samples

Artificial Intelligence Tests

Performance in machine learning and artificial intelligence tasks

GeekBench 6 ML

Tests throughput of AI operations in single, half, and quantized precision
GB6 ML Single Precision
GB6 ML Half Precision
GB6 ML Quantized
Image Classification (SP) - 10338
Image Segmentation (HP) - 42053
Image Super Resolution (Q) - 41729
Face Detection (HP) - 71929
Pose Estimation (Q) - 287035
Text Classification (SP) - 2958
Machine Translation (HP) - 9156
Object Detection (SP) - 9951
Depth Estimation (Q) - 82338
Style Transfer (SP) - 264780
Framework - Core ML
Backend - GPU
Sources: Geekbench [9], [10]

Specifications

Technical specifications of Radeon 8040S and Apple M5 Max GPU (32-core)

General

Vendor Amd Apple
Build Integrated Integrated
Released January 7, 2025 March 3, 2026
Case Laptop Laptop
Purpose Generic Professional
Segment Mid-range Mid-range
Architecture RDNA 3.5 Apple M GPU
GPU Codename Strix Halo Custom
Rival Equivalent - GeForce RTX 3050 Laptop -
Recommended CPU - AMD Ryzen AI Max Pro 380 or above - Apple M5 Max (32-сore GPU) or above
Used in CPUs - AMD Ryzen AI Max PRO 480
- AMD Ryzen AI Max Pro 380
- Apple M5 Max (32-сore GPU)
Laptop GPU ranking (76th and 9th place)

Graphics Processing Unit

Base Clock 1295 MHz -
Boost Clock 2800 MHz 1620 MHz
Shading Units 1024 4096
Texture Mapping Units (TMUs) 64 256
Render Output Units (ROPs) 32 128
Compute Units (Pipelines) 16 512
Tensor Cores No -
Ray-tracing Cores 16 -
L2 Cache 2MB shared -
Instructions Per Cycle 2 IPC 2 IPC

Raw Performance

Pixel Fill Rate 90 GPixel/s 207 GPixel/s
Texture Fill Rate 179 GTexel/s 415 GTexel/s
FLOPS (FP32)
5.7 TFLOPS
13.3 TFLOPS

Physical

Interface PCIe 5.0 x16 Custom
TGP 55 W -
Manufacturing TSMC TSMC
Fabrication Process 4 nm 3 nm
Die Size 308 mm² -
Transistor Count - 28 billion
Max. Temperature 100°C 100°C

Memory

Memory Type System Shared System Shared
Memory Clock - 9600 MHz
Bus - 384-bit
ECC No No
Memory Bandwidth

API

DirectX 12.2 -
Vulkan 1.4 -
OpenGL 4.6 -
OpenCL 2.1 -
CUDA No -
Ray Tracing Yes Yes
DLSS No No
DisplayPort 2.1a -

Cast your vote

Choose between two graphics cards
0 (0%)
0 (0%)
Total votes: < 1

User opinions

You can share your opinion or ask a question in the comments below
🌐 Register your profile and become part of NanoReview community!