NVIDIA GeForce RTX 5090 D vs NVIDIA P102-100 Comparison

NVIDIA
GEFORCE

NVIDIA GeForce RTX 5090 D

CORE STATE GB202
VRAM 32 GB
CLOCK SPEED 2407 MHz
TDP 575 W
BUS WIDTH 512 bit
ARCHITECTURE Blackwell 2.0
nm
PROCESS 5 nm
LAUNCH DATE 2025
VS
NVIDIA
GEFORCE

P102-100

CORE STATE GP102
VRAM 5 GB
CLOCK SPEED 1683 MHz
TDP 250 W
BUS WIDTH 320 bit
ARCHITECTURE Pascal
nm
PROCESS 16 nm
LAUNCH DATE 2018

PERFORMANCE BENCHMARKS

3dmark_3dmark_steel_nomad_dx12
14,326
N/A
geekbench_opencl
310,674
49,602
geekbench_vulkan
376,915
67,454
passmark_directx_10
231
N/A
passmark_directx_11
371
N/A
passmark_directx_12
219
N/A
passmark_directx_9
434
N/A
passmark_g2d
1,487
N/A
passmark_g3d
44,065
N/A
passmark_gpu_compute
28,396
N/A

Analysis: NVIDIA GeForce RTX 5090 D vs NVIDIA P102-100

Head-to-Head Benchmarks

The recorded data shows a decisive outcome in every shared test. The NVIDIA GeForce RTX 5090 D wins both head-to-head benchmark comparisons against the NVIDIA P102-100, with margins that are not close. In Geekbench OpenCL, the RTX 5090 D scores 310,674 against 49,602 for the P102-100, a delta of 526.3%. In Geekbench Vulkan, the RTX 5090 D records 376,915 versus 67,454, a delta of 458.8%. These are not incremental gains; they represent a generational leap in compute performance.

The RTX 5090 D also holds a commanding position in the broader database rankings. Its average benchmark score of 77,712 places it in the 92nd percentile of all GPUs tracked. Its nearest rivals in the database are the AMD Radeon RX 6650M XT at 76,904 (1.1% behind), the AMD Radeon RX 6850M XT at 78,940 (1.6% ahead), and the NVIDIA Tesla P100 variants at 79,396 and 79,605 (2.1% and 2.4% ahead respectively). The RTX 5090 D sits within a tight cluster of high-end mobile and data center parts, but its raw score in the head-to-head tests is what separates it from the P102-100.

The P102-100, by contrast, averages 58,528 across its recorded benchmarks, putting it in the 88th percentile. Its nearest rivals are the AMD Radeon PRO V710 at 58,657 (0.2% ahead), the AMD Radeon RX 6950 XT at 58,392 (0.2% behind), the Intel Arc A570M at 58,239 (0.5% behind), and the AMD Radeon RX 5600 OEM at 58,085 (0.8% behind). The P102-100 is competitive within its own performance tier, but that tier is far below the RTX 5090 D.

Looking at the broader benchmark suite for the RTX 5090 D, the data reveals consistent strength across multiple APIs. In 3DMark Steel Nomad DX12, it scores 14,326. In PassMark tests, it records 44,065 in G3D, 28,396 in GPU Compute, 1,487 in G2D, and legacy DirectX scores of 434 (DX9), 371 (DX11), 231 (DX10), and 219 (DX12). The P102-100 has no recorded scores in these tests, so the comparison is limited to the two Geekbench workloads. Within those, the RTX 5090 D is more than five times faster in OpenCL and more than four and a half times faster in Vulkan.

The magnitude of the win is consistent with the architectural gap between the two parts. The RTX 5090 D is built on the Blackwell 2.0 architecture with a 5 nm process, while the P102-100 uses the Pascal architecture on a 16 nm process. The RTX 5090 D delivers 104.8 TFLOPS of FP32 performance and 104.8 TFLOPS of FP16, while the P102-100 delivers 10.77 TFLOPS FP32 and only 168.3 GFLOPS FP16. The RTX 5090 D also has a massive memory advantage: 32 GB of GDDR7 on a 512-bit bus with 1.79 TB/s bandwidth, versus 5 GB of GDDR5X on a 320-bit bus with 440.3 GB/s.

The Verdict

The verdict is unambiguous: the NVIDIA GeForce RTX 5090 D is the superior product in every measured category. It wins both head-to-head benchmarks, holds a higher average score, ranks in a higher percentile, and offers substantially more compute, memory, and bandwidth. The P102-100 is an end-of-life mining GPU with no display outputs, a PCIe 1.0 x4 interface, and a much older architecture. For any workload captured in the database, the RTX 5090 D is the only rational choice.

The data does not support any scenario where the P102-100 comes out ahead. Its only recorded wins are relative to its own peer group, where it sits within 1% of rivals like the AMD Radeon PRO V710 and AMD Radeon RX 6950 XT. But against the RTX 5090 D, the gap is enormous. In OpenCL, the RTX 5090 D is 526.3% faster. In Vulkan, it is 458.8% faster. These are the largest deltas in the head-to-head data, and they leave no room for interpretation.

The RTX 5090 D also carries a launch MSRP of 2,299 USD, noted once here for reference. The P102-100 has no recorded launch MSRP, which further limits its appeal as a comparison point. The RTX 5090 D is an active product with a successor planned (GeForce 60), while the P102-100 is end-of-life with no successor.

For buyers or analysts evaluating these two GPUs, the database points firmly to the RTX 5090 D. It is faster, newer, more capable, and still in production. The P102-100 should be considered only as a historical or niche part, not as a competitor.

Where Each One Wins

The RTX 5090 D wins in every benchmark category where both are tested. That includes OpenCL compute workloads, where its 310,674 score dwarfs the P102-100's 49,602. It also wins in Vulkan graphics and compute, with 376,915 against 67,454. Beyond the shared tests, the RTX 5090 D has recorded scores in 3DMark Steel Nomad DX12 and the full PassMark suite, giving it a much broader performance profile.

The P102-100 has no recorded wins. Its nearest rivals in the database are all within 1% of its average score, which suggests it is a competent mid-range part for its era. But it lacks ray tracing cores, tensor cores, and any display outputs, which limits its use cases to compute or mining workloads. Its 5 GB memory capacity and 440.3 GB/s bandwidth are sufficient for some tasks, but they are far below the RTX 5090 D's 32 GB and 1.79 TB/s.

The RTX 5090 D also wins on architectural features. It has 21,760 shading units, 680 TMUs, 176 ROPs, 170 RT cores, and 680 tensor cores. The P102-100 has 3,200 shading units, 200 TMUs, and 80 ROPs, with no RT or tensor cores at all. The RTX 5090 D supports DirectX 12 Ultimate (12_2), while the P102-100 only reaches DirectX 12 (12_1). Both support OpenGL 4.6 and Vulkan 1.4, but the RTX 5090 D adds HDMI 2.1b and DisplayPort 2.1b outputs, while the P102-100 has none.

For use cases, the RTX 5090 D is suited to modern gaming, ray tracing, AI inference, and high-bandwidth compute. The P102-100 is limited to tasks that do not require display output or modern API features. The data shows no scenario where the P102-100 is preferable.

FAQ

Q: How much faster is the RTX 5090 D than the P102-100 in OpenCL?

A: The RTX 5090 D scores 310,674 in Geekbench OpenCL, while the P102-100 scores 49,602. That is a 526.3% advantage for the RTX 5090 D.

Q: What is the average benchmark score for each GPU?

A: The RTX 5090 D has an average benchmark score of 77,712, placing it in the 92nd percentile. The P102-100 averages 58,528, placing it in the 88th percentile.

Q: Does the P102-100 support ray tracing or tensor cores?

A: No. The P102-100 has no RT cores and no tensor cores. The RTX 5090 D has 170 RT cores and 680 tensor cores.

Q: What memory configurations do the two GPUs use?

A: The RTX 5090 D uses 32 GB of GDDR7 on a 512-bit bus with 1.79 TB/s bandwidth. The P102-100 uses 5 GB of GDDR5X on a 320-bit bus with 440.3 GB/s bandwidth.

Q: Are there any display outputs on the P102-100?

A: No. The P102-100 has no display outputs. The RTX 5090 D offers 1x HDMI 2.1b and 3x DisplayPort 2.1b.

Q: What are the closest rivals to the RTX 5090 D in the database?

A: The nearest rivals are the AMD Radeon RX 6650M XT at 76,904 (1.1% behind), the AMD Radeon RX 6850M XT at 78,940 (1.6% ahead), the NVIDIA Tesla P100 PCIe 12 GB at 79,396 (2.1% ahead), and the NVIDIA Tesla P100 PCIe 16 GB at 79,605 (2.4% behind).

Architecture Differences

The architectural divide between the RTX 5090 D and the P102-100 is stark. The RTX 5090 D uses the GB202 chip built on the Blackwell 2.0 architecture at TSMC's 5 nm process. It contains 92,200 million transistors on a die size of 750 mm², giving a transistor density of 122.9M per mm². The P102-100 uses the GP102 chip on the Pascal architecture on TSMC's 16 nm process. It has 11,800 million transistors on a 471 mm² die size, with a density of 25.1M per mm².

The RTX 5090 D packs 21,760 shading units, 680 TMUs, 176 ROPs, 170 RT cores, and 680 tensor cores. The P102-100 has 3200 shading units, 200 TMUs, 80 ROPs, and no RT or tensor cores. The RTX 5090 D reaches a pixel rate of 423.6 GPixel/s and a texture rate of 1,636.8 GTexel/s. The P102-100 reaches 134.6 GPixel/s and 336.6 GTexel/s respectively.

The RTX 5090 D supports DirectX 12 Ultimate (12_2), OpenGL 4.6, and Vulkan 1.4. The P102-100 supports DirectX 100, OpenGL 4.6, and Vulkan 1.4. The RTX 5090 D also has a 12_1 bus interface, while the P102-100's is 1.0 x4, a generation behind.

The RTX 5090 D uses GDDR7 memory at 1750 MHz (28 Gbps effective) across a 512-bit bus. The P102-100 uses GDDR5X at 1376 MHz (11 Gbps effective) across a 320-bit bus. The RTX 5090 D is built for compute-heavy workloads, while the P102-100 is a mining-focused part with no outputs. The RTX 5090 D's process node is 5 nm, its predecessor is the GeForce 40 series. The P102-100 has no predecessor, successor, or launch MSRP recorded.

Specification Differences

The specification differences between the two GPUs are extensive. The RTX 5090 D has a base clock of 2017 MHz and a boost clock of 2407 MHz. The P102-100 has a base clock of 1582 MHz and a boost clock of 1683 MHz. The RTX 5090 D has a TDP of 575 W, dual-slot width, and a 1x 16-pin power connector with a suggested PSU of 950 W. The P102-100 has a TDP of 575 W, dual-slot width, 2x 8-pin connectors, and a suggested PSU of 600 W.

The RTX 5090 D supports 32 GB of GDDR7 memory, a 512-bit bus, 1.79 TB/s bandwidth, 21,760 shading units, 680 TMUs, 176 ROPs, 170 RT cores, 680 tensor cores, 423.6 GPixel/s pixel rate 1,636.8 GTexel/s texture rate, 104.8 TFLOPS FP32, 104.8 TFLOPS FP16, 575 W TDP, dual-slot width, 1x 16-pin power connector, 950 W suggested PSU, PCIe 5.0 x16 bus interface, 1x HDMI 2.1b, and 3x DisplayPort 2.1b display outputs, 304 mm length (304 mm 12 inches height, 137 mm 5.4 inches width, 48 mm 1.9 inches width), 5 nm process, 92,200 million transistors, 750 mm² die size, 122.9M / mm² density, 2017 MHz base clock, 2407 MHz boost clock, 1750 MHz memory clock 28 Gbps effective, and a 2025-01-29 release date.

The RTX 5090 D is active production status, with a launch MSRP of 2,299 USD, and a 92nd percentile rank. The P102-100 is 88th percentile, end-of-life, with no launch MSRP. It has 5 GB GDDR5X, 320-bit bus, 440.3 GB/s, 3200 shading units, 200 TMUs, 80 ROPs, 134.6 GPixel/s pixel rate, 336.6 GTexel/s texture rate, 10.77 TFLOPS FP32, 168.3 GFLOPS FP16, 250 W TDP, dual-slot width, 2x 8-pin power connectors, 600 W suggested PSU, PCIe 1.0 x4, no outputs, 267 mm length, 2018-02-11 release date, 16 nm process, 11,800 million transistors, 471 mm² die size, 25.1M / mm² density, 1582 MHz base clock, 1683 MHz boost clock, 1376 MHz memory clock 11 Gbps effective. The RTX has 92.2 billion transistors, 750 mm², 122.9M / mm², while P102 has 11.8 billion, 471 mm², 25.1M / mm². The RTX 5090 D has 104.8 TFLOPS FP32 and FP16 1:1, while P102 has 10.77 TFLOPS FP32 and 168.3 GFLOPS FP16 1:64.

DETAILED SPECIFICATIONS

SPECIFICATION
RTX 5090 D
P102-100
Core Specs
Shading Units
21,760
3,200 -85.3%
Shaders
21,760
3,200 -85.3%
TMUs
680
200 -70.6%
ROPs
176
80 -54.5%
SM Count
170
25 -85.3%
Clocks
Base Clock
2017 MHz
1582 MHz
Boost Clock
2407 MHz
1683 MHz
Memory Clock
1750 MHz 28 Gbps effective
1376 MHz 11 Gbps effective
Memory
Memory Size
32 GB
5 GB
VRAM (MB)
32,768
5,120 -84.4%
Memory Type
GDDR7
GDDR5X
Memory Bus
512 bit
320 bit
Bandwidth
1.79 TB/s
440.3 GB/s
Cache
L1 Cache
128 KB (per SM)
48 KB (per SM)
L2 Cache
96 MB
2.5 MB
Performance
Pixel Rate
423.6 GPixel/s
134.6 GPixel/s
Texture Rate
1,636.8 GTexel/s
336.6 GTexel/s
FP32 (TFLOPS)
104.8 TFLOPS
10.77 TFLOPS
FP64 (TFLOPS)
1.637 TFLOPS (1:64)
336.6 GFLOPS (1:32)
FP16 (TFLOPS)
104.8 TFLOPS (1:1)
168.3 GFLOPS (1:64)
AI/RT
RT Cores
170
Tensor Cores
680
Power
TDP
575 W
250 W
TDP (W)
575
250 -56.5%
Suggested PSU
950 W
600 W
Power Connectors
1x 16-pin
2x 8-pin
Architecture
Architecture
Blackwell 2.0
Pascal
GPU Name
GB202
GP102
Generation
GeForce 50
Mining GPUs
Process Size
5 nm
16 nm
Transistors
92,200 million
11,800 million
Die Size
750 mm²
471 mm²
Foundry
TSMC
TSMC
Density
122.9M / mm²
25.1M / mm²
API Support
DirectX
12 Ultimate (12_2)
12 (12_1)
OpenGL
4.6
4.6
Vulkan
1.4
1.4
OpenCL
3.0
3.0
CUDA
12.0
6.1
Shader Model
6.9
6.8
Physical
Slot Width
Dual-slot
Dual-slot
Length
304 mm 12 inches
267 mm 10.5 inches
Height
137 mm 5.4 inches
Outputs
1x HDMI 2.1b3x DisplayPort 2.1b
No outputs
Bus Interface
PCIe 5.0 x16
PCIe 1.0 x4
Other
Launch Price
2,299 USD
Production
Active
End-of-life
Predecessor
GeForce 40
Successor
GeForce 60
View GeForce RTX 5090 D Details View P102-100 Details