AMD Instinct MI350P vs NVIDIA GeForce RTX 4090 Mobile Comparison

AMD
RADEON

AMD Instinct MI350P

CORE STATE MI350 128CU
VRAM 144 GB
CLOCK SPEED 2200 MHz
TDP 600 W
BUS WIDTH 8192 bit
ARCHITECTURE CDNA 4.0
nm
PROCESS 3 nm
LAUNCH DATE 2026
VS
NVIDIA
GEFORCE

GeForce RTX 4090 Mobile

CORE STATE AD103
VRAM 16 GB
CLOCK SPEED 1695 MHz
TDP 120 W
BUS WIDTH 256 bit
ARCHITECTURE Ada Lovelace
nm
PROCESS 5 nm
LAUNCH DATE 2023

PERFORMANCE BENCHMARKS

geekbench_opencl
N/A
180,831
geekbench_vulkan
N/A
170,774
passmark_directx_10
N/A
173
passmark_directx_11
N/A
262
passmark_directx_12
N/A
107
passmark_directx_9
N/A
310
passmark_g2d
N/A
984
passmark_g3d
N/A
27,212
passmark_gpu_compute
N/A
12,347

Analysis: AMD Instinct MI350P vs NVIDIA GeForce RTX 4090 Mobile

Head-to-Head Benchmarks

The benchmark comparison between the AMD Instinct MI350P and the NVIDIA GeForce RTX 4090 Mobile is limited by the available data. The database records no head-to-head benchmark entries for this pairing, and the AMD Instinct MI350P has no individual benchmark scores listed. The NVIDIA GeForce RTX 4090 Mobile, however, has a full set of recorded measurements across nine tests.

The RTX 4090 Mobile delivers its strongest result in the Passmark G3D test with a score of 27,212. Its compute performance is substantial as well, with a Passmark GPU Compute score of 12,347. In the Geekbench suite, the card records 180,831 in OpenCL and 170,774 in Vulkan. Legacy DirectX tests show scores of 310 for DirectX 9, 262 for DirectX 11, 173 for DirectX 10, and 107 for DirectX 12. The 2D graphics test yields a score of 984.

The RTX 4090 Mobile sits at the 84th percentile among all GPUs in the database, with an average benchmark score of 43,667. Its nearest rivals provide context for this figure. The NVIDIA RTX A6000 posts an average score of 44,075, which is 0.9% higher than the RTX 4090 Mobile. The NVIDIA Quadro M6000 trails by 0.8% with an average score of 43,301. The NVIDIA GeForce RTX 5050 Mobile scores 43,268, a 0.9% deficit, and the NVIDIA Quadro M6000 24 GB scores 43,262, also a 0.9% deficit. These margins place the RTX 4090 Mobile in a tight cluster of high-end workstation and mobile parts.

Because the MI350P has no recorded benchmark scores, no direct performance comparison can be made from the database. The MI350P carries a percentile score of 50, but this figure does not correspond to any measured benchmark result. The data simply does not support a numeric head-to-head comparison between these two accelerators.

Architecture Differences

The AMD Instinct MI350P uses the CDNA 4.0 architecture, built on a 3 nm process at TSMC. It packs 73,000 million transistors on a die size of 1,190 mm², resulting in a transistor density of 61.3 million per mm². The chip is designated MI350 128CU, indicating a compute-oriented design with 128 compute units.

The NVIDIA GeForce RTX 4090 Mobile uses the Ada Lovelace architecture, built on a 5 nm process at TSMC. It contains 45,900 million transistors on a die size of 379 mm², giving it a transistor density of 121.1 million per mm². The chip is designated AD103.

The MI350P is a dual-slot accelerator with a 600 W TDP and a single 16-pin power connector. It uses a PCIe 5.0 x16 bus interface. The RTX 4090 Mobile is an integrated graphics processor (IGP) with a 120 W TDP and no power connectors, using a PCIe 4.0 x16 bus interface.

Memory configurations diverge sharply. The MI350P carries 144 GB of HBM3e memory on an 8,192-bit bus, delivering 8.19 TB/s of bandwidth. The RTX 4090 Mobile has 16 GB of GDDR6 memory on a 256-bit bus, with 576.0 GB/s of bandwidth. The MI350P's memory clock is listed as 2000 MHz with 8 Gbps effective speed, while the RTX 4090 Mobile runs at 2250 MHz with 18 Gbps effective.

The MI350P has 8,192 shading units and 512 texture mapping units, but no ROPs, no RT cores, and no tensor cores. Its pixel rate is listed as 0 MPixel/s, and its texture rate is 1,126.4 GTexel/s. The RTX 4090 Mobile has 9,728 shading units, 304 TMUs, and 112 ROPs. It includes 76 RT cores and 304 tensor cores. Its pixel rate is 189.8 GPixel/s, and its texture rate is 515.3 GTexel/s.

Floating-point performance shows a similar scale. The MI350P delivers 36.04 TFLOPS in both FP32 and FP16 (1:1 ratio). The RTX 4090 Mobile delivers 32.98 TFLOPS in both FP32 and FP16 (1:1 ratio). The MI350P has no display outputs, while the RTX 4090 Mobile's display outputs are listed as portable device dependent.

API support differs completely. The MI350P has no DirectX, OpenGL, or Vulkan support. The RTX 4090 Mobile supports DirectX 12 Ultimate (12_2), OpenGL 4.6, and Vulkan 1.4.

FAQ

Q: Which GPU has higher raw FP32 compute?

A: The AMD Instinct MI350P leads with 36.04 TFLOPS FP32, while the NVIDIA GeForce RTX 4090 Mobile delivers 32.98 TFLOPS FP32.

Q: How do the memory capacities compare?

A: The MI350P has 144 GB of HBM3e memory, which is 9 times the 16 GB of GDDR6 found on the RTX 4090 Mobile.

Q: Does the RTX 4090 Mobile support ray tracing?

A: Yes, the RTX 4090 Mobile includes 76 RT cores and 304 tensor cores, and it supports DirectX 12 Ultimate, OpenGL 4.6, and Vulkan 1.4.

Q: What is the power draw difference?

A: The MI350P has a TDP of 600 W and requires a 1000 W suggested PSU, while the RTX 4090 Mobile has a TDP of 120 W and requires no external power connectors.

Q: Which GPU has a higher transistor density?

A: The RTX 4090 Mobile has a higher density at 121.1 million transistors per mm², compared to the MI350P's 61.3 million per mm².

Q: What is the bus interface difference?

A: The MI350P uses PCIe 5.0 x16, while the RTX 4090 Mobile uses PCIe 4.0 x16.

Specification Differences

| Specification | AMD Instinct MI350P | NVIDIA GeForce RTX 4090 Mobile |

|---|---|---|

| Architecture | CDNA 4.0 | Ada Lovelace |

| Process Node | 3 nm | 5 nm |

| Transistors | 73,000 million | 45,900 million |

| Die Size | 1190 mm² | 379 mm² |

| Transistor Density | 61.3M / mm² | 121.1M / mm² |

| Base Clock | 1000 MHz | 1335 MHz |

| Boost Clock | 2200 MHz | 1695 MHz |

| Memory Size | 144 GB | 16 GB |

| Memory Type | HBM3e | GDDR6 |

| Memory Bus Width | 8192 bit | 256 bit |

| Memory Bandwidth | 8.19 TB/s | 576.0 GB/s |

| Shading Units | 8192 | 9728 |

| TMUs | 512 | 304 |

| ROPs | 0 | 112 |

| RT Cores | null | 76 |

| Tensor Cores | null | 304 |

| Pixel Rate | 0 MPixel/s | 189.8 GPixel/s |

| Texture Rate | 1,126.4 GTexel/s | 515.3 GTexel/s |

| FP32 | 36.04 TFLOPS | 32.98 TFLOPS |

| FP16 | 36.04 TFLOPS (1:1) | 32.98 TFLOPS (1:1) |

| TDP | 600 W | 120 W |

| Slot Width | Dual-slot | IGP |

| Power Connectors | 1x 16-pin | None |

| Suggested PSU | 1000 W | null |

| Bus Interface | PCIe 5.0 x16 | PCIe 4.0 x16 |

| Display Outputs | No outputs | Portable Device Dependent |

| DirectX | N/A | 12 Ultimate (12_2) |

| OpenGL | N/A | 4.6 |

| Vulkan | N/A | 1.4 |

| Release Date | 2026-05-06 | 2023-01-02 |

| Production Status | null | Active |

Where Each One Wins

The AMD Instinct MI350P wins in compute throughput and memory capacity. Its 36.04 TFLOPS FP32 figure exceeds the RTX 4090 Mobile's 32.98 TFLOPS. The 144 GB of HBM3e memory with 8.19 TB/s bandwidth is an order of magnitude larger and faster than the 16 GB GDDR6 at 576.0 GB/s. The texture rate of 1,126.4 GTexel/s more than doubles the RTX 4090 Mobile's 515.3 GTexel/s. The MI350P also uses a newer 3 nm process and a faster PCIe 5.0 interface.

The NVIDIA GeForce RTX 4090 Mobile wins in rasterization, graphics features, and efficiency. It has 112 ROPs versus none on the MI350P, and a pixel rate of 189.8 GPixel/s. The RTX 4090 Mobile includes 76 RT cores and 304 tensor cores, enabling hardware ray tracing and AI acceleration. It supports DirectX 12 Ultimate, OpenGL 4.6, and Vulkan 1.4, while the MI350P has no graphics API support. The RTX 4090 Mobile runs at 120 W TDP, a fraction of the MI350P's 600 W, and requires no external power connectors. Its 9,728 shading units outnumber the MI350P's 8,192. The RTX 4090 Mobile also has a higher transistor density at 121.1M per mm² versus 61.3M per mm².

The RTX 4090 Mobile is the only one with recorded benchmark results. Its average score of 43,667 places it at the 84th percentile, with its nearest rivals all within 0.9% of its score. The MI350P has no benchmark data, so its practical performance in the database remains unmeasured.

The Verdict

The data describes two accelerators built for entirely different tasks. The AMD Instinct MI350P is a compute-oriented accelerator with no display outputs and no graphics API support. It prioritizes raw throughput, memory capacity, and bandwidth for compute workloads. The 144 GB HBM3e pool and 8.19 TB/s bandwidth are the standout features, along with the 36.04 TFLOPS FP32 figure. Its 600 W TDP and dual-slot form factor indicate a server or workstation installation, not a desktop or laptop.

The NVIDIA GeForce RTX 4090 Mobile is a laptop-class GPU with full graphics capabilities. It supports DirectX 12 Ultimate, OpenGL 4.6, and Vulkan 1.4, includes RT and tensor cores, and delivers a measured average benchmark score of 43,667. Its 120 W TDP and IGP form factor make it suitable for portable devices. The benchmark scores show a balanced performer: the Passmark G3D score of 27,212 and Geekbench OpenCL score of 180,831 confirm solid rasterization and compute capability.

For users requiring maximum memory and compute density in a server context, the MI350P is the clear choice based on its specifications. For any workload involving graphics rendering, ray tracing, or portable deployment, the RTX 4090 Mobile is the only option with measured performance and software support. The database does not contain evidence of a direct performance comparison, so a definitive winner cannot be declared. The selection depends entirely on the workload and form factor requirements.

DETAILED SPECIFICATIONS

SPECIFICATION
Instinct MI350P
RTX 4090 Mobile
Core Specs
Shading Units
8,192
9,728 +18.8%
Shaders
8,192
9,728 +18.8%
TMUs
512
304 -40.6%
ROPs
0
112 +∞%
Compute Units
128
SM Count
76
Clocks
Base Clock
1000 MHz
1335 MHz
Boost Clock
2200 MHz
1695 MHz
Memory Clock
2000 MHz 8 Gbps effective
2250 MHz 18 Gbps effective
Memory
Memory Size
144 GB
16 GB
VRAM (MB)
147,456
16,384 -88.9%
Memory Type
HBM3e
GDDR6
Memory Bus
8192 bit
256 bit
Bandwidth
8.19 TB/s
576.0 GB/s
Cache
L1 Cache
16 KB (per CU)
128 KB (per SM)
L2 Cache
16 MB
64 MB
L3 Cache
128 MB
Performance
Pixel Rate
0 MPixel/s
189.8 GPixel/s
Texture Rate
1,126.4 GTexel/s
515.3 GTexel/s
FP32 (TFLOPS)
36.04 TFLOPS
32.98 TFLOPS
FP64 (TFLOPS)
18.02 TFLOPS (1:2)
515.3 GFLOPS (1:64)
FP16 (TFLOPS)
36.04 TFLOPS (1:1)
32.98 TFLOPS (1:1)
AI/RT
RT Cores
76
Tensor Cores
304
Matrix Cores
512
Power
TDP
600 W
120 W
TDP (W)
600
120 -80.0%
Suggested PSU
1000 W
Power Connectors
1x 16-pin
None
Architecture
Architecture
CDNA 4.0
Ada Lovelace
GPU Name
MI350 128CU
AD103
Generation
Instinct (MIx)
GeForce 40 Mobile
Process Size
3 nm
5 nm
Transistors
73,000 million
45,900 million
Die Size
1190 mm²
379 mm²
Foundry
TSMC
TSMC
Density
61.3M / mm²
121.1M / mm²
AMD MCM
MCM
2
API Support
DirectX
12 Ultimate (12_2)
OpenGL
4.6
Vulkan
1.4
OpenCL
3.0
3.0
CUDA
8.9
Shader Model
6.8
Physical
Slot Width
Dual-slot
IGP
Length
267 mm 10.5 inches
Height
111 mm 4.4 inches
Outputs
No outputs
Portable Device Dependent
Bus Interface
PCIe 5.0 x16
PCIe 4.0 x16
Other
Production
Active
Predecessor
Radeon Instinct
GeForce 30 Mobile
Successor
GeForce 50 Mobile
View Instinct MI350P Details View GeForce RTX 4090 Mobile Details