GPU Comparison
RADEON
AMD Instinct MI100
CORE STATE
Arcturus
VRAM
32 GB
CLOCK SPEED
1502 MHz
TDP
300 W
BUS WIDTH
4096 bit
ARCHITECTURE
CDNA 1.0
PROCESS
7 nm
LAUNCH DATE
2020
VS
GEFORCE
L40S
CORE STATE
AD102
VRAM
48 GB
CLOCK SPEED
2520 MHz
TDP
300 W
BUS WIDTH
384 bit
ARCHITECTURE
Ada Lovelace
PROCESS
5 nm
LAUNCH DATE
2022
PERFORMANCE BENCHMARKS
geekbench_opencl
geekbench_vulkan
DETAILED SPECIFICATIONS
SPECIFICATION
Instinct MI100
L40S
Core Specs
Shading Units
7,680
18,176
+136.7%
Shaders
7,680
18,176
+136.7%
TMUs
480
568
+18.3%
ROPs
64
192
+200.0%
Compute Units
120
—
SM Count
—
142
Clocks
Base Clock
1000 MHz
1110 MHz
Boost Clock
1502 MHz
2520 MHz
Memory Clock
1200 MHz 2.4 Gbps effective
2250 MHz
18 Gbps effective
Memory
Memory Size
32 GB
48 GB
VRAM (MB)
32,768
49,152
+50.0%
Memory Type
HBM2
GDDR6
Memory Bus
4096 bit
384 bit
Bandwidth
1.23 TB/s
864.0 GB/s
Cache
L1 Cache
16 KB (per CU)
128 KB (per SM)
L2 Cache
8 MB
48 MB
Performance
Pixel Rate
96.13 GPixel/s
483.8 GPixel/s
Texture Rate
721.0 GTexel/s
1,431.4 GTexel/s
FP32 (TFLOPS)
23.07 TFLOPS
91.61 TFLOPS
FP64 (TFLOPS)
11.54 TFLOPS (1:2)
1,431.4 GFLOPS (1:64)
FP16 (TFLOPS)
46.14 TFLOPS (2:1)
91.61 TFLOPS (1:1)
AI/RT
RT Cores
—
142
Tensor Cores
—
568
Power
TDP
300 W
300 W
TDP (W)
300
300
0.0%
Suggested PSU
700 W
700 W
Power Connectors
2x 8-pin
1x 16-pin
Architecture
Architecture
CDNA 1.0
Ada Lovelace
GPU Name
Arcturus
AD102
Generation
Instinct (MIx)
Server Ada
(Lxx)
Process Size
7 nm
5 nm
Transistors
25,600 million
76,300 million
Die Size
750 mm²
609 mm²
Foundry
TSMC
TSMC
Density
34.1M / mm²
125.3M / mm²
API Support
DirectX
—
12 Ultimate (12_2)
OpenGL
—
4.6
Vulkan
—
1.4
OpenCL
2.1
3.0
CUDA
—
8.9
Shader Model
—
6.8
Physical
Slot Width
Dual-slot
Dual-slot
Length
267 mm 10.5 inches
267 mm
10.5 inches
Height
111 mm 4.4 inches
111 mm
4.4 inches
Outputs
No outputs
1x HDMI 2.13x DisplayPort 1.4a
Bus Interface
PCIe 4.0 x16
PCIe 4.0 x16
Other
Production
End-of-life
End-of-life
Predecessor
Radeon Instinct
Server Ampere
Successor
—
Server Hopper