AMD Instinct MI325X vs NVIDIA GeForce RTX 5090 D Comparison

AMD
RADEON

AMD Instinct MI325X

CORE STATE Aqua Vanjaram
VRAM 256 GB
CLOCK SPEED 2100 MHz
TDP 1000 W
BUS WIDTH 8192 bit
ARCHITECTURE CDNA 3.0
nm
PROCESS 5 nm
LAUNCH DATE 2024
VS
NVIDIA
GEFORCE

GeForce RTX 5090 D

CORE STATE GB202
VRAM 32 GB
CLOCK SPEED 2407 MHz
TDP 575 W
BUS WIDTH 512 bit
ARCHITECTURE Blackwell 2.0
nm
PROCESS 5 nm
LAUNCH DATE 2025

PERFORMANCE BENCHMARKS

3dmark_3dmark_steel_nomad_dx12
N/A
14,326
geekbench_opencl
N/A
310,674
geekbench_vulkan
N/A
376,915
passmark_directx_10
N/A
231
passmark_directx_11
N/A
371
passmark_directx_12
N/A
219
passmark_directx_9
N/A
434
passmark_g2d
N/A
1,487
passmark_g3d
N/A
44,065
passmark_gpu_compute
N/A
28,396

Analysis: AMD Instinct MI325X vs NVIDIA GeForce RTX 5090 D

The Verdict

The database shows two accelerators built for entirely different workloads. The AMD Instinct MI325X is a compute-focused accelerator with no display outputs, no graphics API support, and a massive 256 GB HBM3e memory pool. The NVIDIA GeForce RTX 5090 D is a consumer graphics card with full DirectX 12 Ultimate, OpenGL 4.6, and Vulkan 1.4 support, plus display outputs. The recorded benchmark data covers only the RTX 5090 D, which holds a 92nd percentile ranking across all GPUs with an average benchmark score of 77,712. The MI325X has no recorded benchmark scores and sits at the 50th percentile with an average score of zero, so direct performance comparisons in this database are limited to architectural and specification analysis rather than measured head-to-head results.

The RTX 5090 D delivers higher FP32 throughput at 104.8 TFLOPS compared to 81.72 TFLOPS for the MI325X, and it is the only one of the two with rasterization hardware, featuring 176 ROPs and a pixel rate of 423.6 GPixel/s. The MI325X instead prioritizes memory capacity and bandwidth, offering 256 GB of HBM3e across an 8192-bit bus for 6.14 TB/s of bandwidth, which is roughly 3.4 times the bandwidth of the RTX 5090 D's 1.79 TB/s over a 512-bit GDDR7 interface. Buyers needing graphics output, ray tracing, or consumer software support should select the RTX 5090 D. Buyers needing maximum memory capacity for large datasets should select the MI325X.

Architecture Differences

The MI325X uses the Aqua Vanjaram chip built on CDNA 3.0 architecture, fabricated on a 5 nm process at TSMC. The RTX 5090 D uses the GB202 chip built on Blackwell 2.0 architecture, also on a 5 nm TSMC process. The MI325X packs 153,000 million transistors into a 1017 mm² die, yielding a transistor density of 150.4 million per mm². The RTX 5090 D contains 92,200 million transistors on a 750 mm² die, for a density of 122.9 million per mm². The MI325X has the larger die and the higher transistor count, while the RTX 5090 D achieves higher clock speeds: its base clock is 2017 MHz and boost clock is 2407 MHz, versus 1000 MHz base and 2100 MHz boost for the MI325X.

The MI325X has 19,456 shading units and 1,216 texture mapping units, but zero ROPs, zero ray tracing cores, and zero tensor cores. Its pixel rate is listed as 0 MPixel/s. The RTX 5090 D has 21,760 shading units, 680 TMUs, 176 ROPs, 170 ray tracing cores, and 680 tensor cores. The MI325X texture rate is 2,553.6 GTexel/s, which exceeds the RTX 5090 D's 1,636.8 GTexel/s despite the NVIDIA card having more shading units. The RTX 5090 D supports DirectX 12 Ultimate (12_2), OpenGL 4.6, and Vulkan 1.4; the MI325X lists N/A for all three graphics APIs. The MI325X has no display outputs, while the RTX 5090 D provides 1x HDMI 2.1b and 3x DisplayPort 2.1b.

Head-to-Head Benchmarks

The database contains no shared head-to-head benchmark entries between the two products. The MI325X has an empty benchmarks array, an average benchmark score of zero, and no nearest rivals. The RTX 5090 D has ten recorded benchmark scores. Its 3DMark Steel Nomad DX12 score is 14,326. In Geekbench, it scores 310,674 in OpenCL and 376,915 in Vulkan. Passmark results include 231 in DirectX 10, 371 in DirectX 11, 219 in DirectX 12, 434 in DirectX 9, 1,487 in G2D, 44,065 in G3D, and 28,396 in GPU compute.

The nearest rivals for the RTX 5090 D in the database are AMD Radeon RX 6650M XT with an average score of 76,904 and a delta of +1.1%, AMD Radeon RX 6850M XT at 78,940 with a delta of -1.6%, NVIDIA Tesla P100 PCIe 12 GB at 79,396 with a delta of -2.1%, and NVIDIA Tesla P100 PCIe 16 GB at 79,605 with a delta of -2.4%. These deltas indicate the RTX 5090 D's average benchmark score of 77,712 sits within roughly 2.4% of these four rivals, meaning its measured performance is closely clustered with those products despite its high percentile ranking. The RTX 5090 D's 92nd percentile placement reflects its position above the majority of all GPUs in the database, but its nearest rivals show that the immediate competitive field is tight.

Specification Differences

The two products differ across nearly every specification field. The MI325X uses HBM3e memory with 256 GB capacity, an 8192-bit bus, and 6.14 TB/s bandwidth. The RTX 5090 D uses GDDR7 memory with 32 GB capacity, a 512-bit bus, and 1.79 TB/s bandwidth. Memory clock differs substantially: the MI325X runs at 1500 MHz with 6 Gbps effective, while the RTX 5090 D runs at 1750 MHz with 28 Gbps effective. The MI325X has more TMUs at 1,216 versus 680, but the RTX 5090 D has more shading units at 21,760 versus 19,456. The MI325X has zero ROPs; the RTX 5090 D has 176. The RTX 5090 D has 170 ray tracing cores and 680 tensor cores; the MI325X has none of either.

Power draw differs by a wide margin. The MI325X has a TDP of 1000 W with a suggested PSU of 1400 W, uses an OAM Module slot width, and has no power connectors listed. The RTX 5090 D has a 575 W TDP with a suggested PSU of 950 W, is dual-slot, and uses a single 16-pin connector. Both use PCIe 5.0 x16. The RTX 5090 D measures 304 mm in length, 137 mm in height, and 48 mm in width. The MI325X has no listed dimensions. The MI325X was released on 2024-10-09 and has no production status listed; the RTX 5090 D was released on 2025-01-29 and is marked as Active production. The RTX 5090 D has a predecessor of GeForce 40 and a successor of GeForce 60; the MI325X predecessor is Radeon Instinct and it has no successor listed. The RTX 5090 D launch MSRP is 2,299 USD.

FAQ

Q: Which card has more memory bandwidth?

A: The AMD Instinct MI325X provides 6.14 TB/s of bandwidth from 256 GB of HBM3e on an 8192-bit bus, compared to 1.79 TB/s from 32 GB of GDDR7 on a 512-bit bus for the NVIDIA GeForce RTX 5090 D.

Q: Does the MI325X support graphics APIs?

A: No. The MI325X lists DirectX, OpenGL, and Vulkan as N/A and has no display outputs. The RTX 5090 D supports DirectX 12 Ultimate (12_2), OpenGL 4.6, and Vulkan 1.4, and provides 1x HDMI 2.1b and 3x DisplayPort 2.1b.

Q: Which product has the higher FP32 compute throughput?

A: The NVIDIA GeForce RTX 5090 D delivers 104.8 TFLOPS FP32, while the AMD Instinct MI325X delivers 81.72 TFLOPS FP32. Both list FP16 at a 1:1 ratio with their FP32 figures.

Q: How does the RTX 5090 D compare to its nearest rivals in the database?

A: Its average benchmark score is 77,712. It sits 1.1% above the AMD Radeon RX 6650M XT (76,904), 1.6% below the AMD Radeon RX 6850M XT (78,940), 2.1% below the NVIDIA Tesla P100 PCIe 12 GB (79,396), and 2.4% below the NVIDIA Tesla P100 PCIe 16 GB (79,605).

Q: What are the transistor counts of each chip?

A: The MI325X's Aqua Vanjaram chip contains 153,000 million transistors on a 1017 mm² die. The RTX 5090 D's GB202 chip contains 92,200 million transistors on a 750 mm² die.

Q: Which product requires more power?

A: The MI325X has a 1000 W TDP and a suggested PSU of 1400 W. The RTX 5090 D has a 575 W TDP and a suggested PSU of 950 W.

Where Each One Wins

The RTX 5090 D wins in graphics-centric workloads. It is the only product with ROPs, ray tracing cores, tensor cores, and a non-zero pixel rate of 423.6 GPixel/s. It supports DirectX 12 Ultimate, OpenGL 4.6, and Vulkan 1.4, and it has display outputs. Its 104.8 TFLOPS FP32 exceeds the MI325X by roughly 28%. Its base clock of 2017 MHz is more than double the MI325X's 1000 MHz base, and its boost clock of 2407 MHz is higher as well. It has more shading units at 21,760 versus 19,456, and its 176 ROPs enable rasterization that the MI325X simply cannot perform. The RTX 5090 D is also the only one with recorded benchmark scores, holding a 92nd percentile ranking across all GPUs.

The MI325X wins in memory capacity and bandwidth. Its 256 GB HBM3e pool is eight times the capacity of the RTX 5090 D's 32 GB. Its 6.14 TB/s bandwidth is roughly 3.4 times the RTX 5090 D's 1.79 TB/s. The MI325X also has a wider memory bus at 8192 bits versus 512 bits. Its texture rate of 2,553.6 GTexel/s is higher than the RTX 5090 D's 1,636.8 GTexel/s, and it has more TMUs at 1,216 versus 680. The MI325X uses an OAM Module form factor with no power connectors, indicating a server-oriented design, and it carries more transistors at 153,000 million compared to 92,200 million, on a larger 1017 mm² die versus 750 mm².

The data indicates a clear split. The RTX 5090 D is the choice for graphics, ray tracing, and consumer software environments. The MI325X is the choice for memory-bound compute tasks where 256 GB of HBM3e and 6.14 TB/s of bandwidth matter more than rasterization or graphics API support. The RTX 5090 D holds the only recorded benchmark results in this database, so its measured performance is verifiable, while the MI325X's capabilities are documented only through its specifications.

DETAILED SPECIFICATIONS

SPECIFICATION
Instinct MI325X
RTX 5090 D
Core Specs
Shading Units
19,456
21,760 +11.8%
Shaders
19,456
21,760 +11.8%
TMUs
1,216
680 -44.1%
ROPs
0
176 +∞%
Compute Units
304
—
SM Count
—
170
Clocks
Base Clock
1000 MHz
2017 MHz
Boost Clock
2100 MHz
2407 MHz
Memory Clock
1500 MHz 6 Gbps effective
1750 MHz 28 Gbps effective
Memory
Memory Size
256 GB
32 GB
VRAM (MB)
262,144
32,768 -87.5%
Memory Type
HBM3e
GDDR7
Memory Bus
8192 bit
512 bit
Bandwidth
6.14 TB/s
1.79 TB/s
Cache
L1 Cache
16 KB (per CU)
128 KB (per SM)
L2 Cache
16 MB
96 MB
L3 Cache
256 MB
—
Performance
Pixel Rate
0 MPixel/s
423.6 GPixel/s
Texture Rate
2,553.6 GTexel/s
1,636.8 GTexel/s
FP32 (TFLOPS)
81.72 TFLOPS
104.8 TFLOPS
FP64 (TFLOPS)
40.86 TFLOPS (1:2)
1.637 TFLOPS (1:64)
FP16 (TFLOPS)
81.72 TFLOPS (1:1)
104.8 TFLOPS (1:1)
AI/RT
RT Cores
—
170
Tensor Cores
—
680
Matrix Cores
1,216
—
Power
TDP
1000 W
575 W
TDP (W)
1,000
575 -42.5%
Suggested PSU
1400 W
950 W
Power Connectors
None
1x 16-pin
Architecture
Architecture
CDNA 3.0
Blackwell 2.0
GPU Name
Aqua Vanjaram
GB202
Generation
Instinct (MIx)
GeForce 50
Process Size
5 nm
5 nm
Transistors
153,000 million
92,200 million
Die Size
1017 mm²
750 mm²
Foundry
TSMC
TSMC
Density
150.4M / mm²
122.9M / mm²
AMD MCM
MCM
2
—
API Support
DirectX
—
12 Ultimate (12_2)
OpenGL
—
4.6
Vulkan
—
1.4
OpenCL
3.0
3.0
CUDA
—
12.0
Shader Model
—
6.9
Physical
Slot Width
OAM Module
Dual-slot
Length
—
304 mm 12 inches
Height
—
137 mm 5.4 inches
Outputs
No outputs
1x HDMI 2.1b3x DisplayPort 2.1b
Bus Interface
PCIe 5.0 x16
PCIe 5.0 x16
Other
Launch Price
—
2,299 USD
Production
—
Active
Predecessor
Radeon Instinct
GeForce 40
Successor
—
GeForce 60
View Instinct MI325X Details View GeForce RTX 5090 D Details