NVIDIA GeForce RTX 4070 Mobile vs NVIDIA Rubin GPU Comparison

NVIDIA
GEFORCE

NVIDIA GeForce RTX 4070 Mobile

CORE STATE AD106
VRAM 8 GB
CLOCK SPEED 1695 MHz
TDP 115 W
BUS WIDTH 128 bit
ARCHITECTURE Ada Lovelace
nm
PROCESS 5 nm
LAUNCH DATE 2023
VS
NVIDIA
GEFORCE

Rubin GPU

CORE STATE GR100
VRAM 288 GB
CLOCK SPEED 2267 MHz
TDP 2300 W
BUS WIDTH 16384 bit
ARCHITECTURE Rubin
nm
PROCESS 3 nm
LAUNCH DATE 2026

PERFORMANCE BENCHMARKS

geekbench_opencl
109,197
N/A
geekbench_vulkan
108,367
N/A
passmark_directx_10
116
N/A
passmark_directx_11
179
N/A
passmark_directx_12
85
N/A
passmark_directx_9
223
N/A
passmark_g2d
763
N/A
passmark_g3d
19,587
N/A
passmark_gpu_compute
8,399
N/A

Analysis: NVIDIA GeForce RTX 4070 Mobile vs NVIDIA Rubin GPU

FAQ

Q: How does the NVIDIA GeForce RTX 4070 Mobile compare to the NVIDIA Rubin GPU in terms of average benchmark score?

A: The RTX 4070 Mobile has an average benchmark score of 27435, while the Rubin GPU has no recorded benchmark scores and an average of 0. The database places the RTX 4070 Mobile in the 73rd percentile of all GPUs, whereas the Rubin GPU sits at the 50th percentile with no test data.

Q: What are the closest rivals to the RTX 4070 Mobile in the database?

A: The nearest rivals are the AMD Radeon RX 6700 XT with an average score of 27425 and a 0% delta, the NVIDIA GeForce RTX 3090 with 27565 and a -0.5% delta, the NVIDIA RTX PRO 4000 Blackwell with 27135 and a +1.1% delta, and the AMD Radeon Pro Vega 20 with 27839 and a -1.5% delta.

Q: What memory configurations do the two GPUs use?

A: The RTX 4070 Mobile uses 8 GB of GDDR6 memory on a 128-bit bus with 256.0 GB/s bandwidth. The Rubin GPU uses 288 GB of HBM4 memory on a 16384-bit bus with 22.1 TB/s bandwidth.

Q: What process nodes are used by each GPU?

A: The RTX 4070 Mobile is built on a 5 nm process at TSMC, while the Rubin GPU is built on a 3 nm process, also at TSMC.

Q: What are the transistor counts and die sizes?

A: The RTX 4070 Mobile has 22,900 million transistors on a 188 mm² die, giving a density of 121.8M per mm². The Rubin GPU has 336,000 million transistors on a 1456 mm² die, with a density of 230.8M per mm².

Q: Does the Rubin GPU have display outputs?

A: No, the Rubin GPU has no display outputs, while the RTX 4070 Mobile's display outputs are portable device dependent. The Rubin GPU is a server SXM module, whereas the RTX 4070 Mobile is an integrated graphics processor (IGP).

Architecture Differences

The RTX 4070 Mobile uses the AD106 chip based on the Ada Lovelace architecture, part of the GeForce 40 Mobile generation. The Rubin GPU uses the GR100 chip based on the Rubin architecture, part of the Server Rubin (Rxx) generation. These are fundamentally different design targets: one for mobile graphics, the other for server compute.

The process nodes differ significantly. The RTX 4070 Mobile uses a 5 nm TSMC process, while the Rubin GPU uses a 3 nm TSMC process. This accompanies a dramatic difference in scale: the Rubin GPU packs 336,000 million transistors across a 1456 mm² die, versus 22,900 million transistors on 188 mm² for the mobile part. Transistor density rises from 121.8M per mm² to 230.8M per mm².

Shading unit counts differ by a factor of over six. The RTX 4070 Mobile has 4608 shading units, 144 TMUs, and 48 ROPs. The Rubin GPU has 28672 shading units, 896 TMUs, and only 24 ROPs. The low ROP count on the Rubin part reflects its compute-oriented role rather than rasterization focus. Ray tracing cores are present on the mobile GPU with 36, but the Rubin GPU lists no ray tracing core count. Tensor cores scale from 144 on the mobile part to 896 on the Rubin part.

Clock behavior differs sharply. The RTX 4070 Mobile runs at a base of 1395 MHz and boost of 1695 MHz. The Rubin GPU has a much lower base clock of 700 MHz but a higher boost of 2267 MHz. Memory clocks also diverge: the mobile part uses 2000 MHz with 16 Gbps effective, while the Rubin GPU uses 2695 MHz with 10.8 Gbps effective.

API support separates the two products. The RTX 4070 Mobile supports DirectX 12 Ultimate (12_2), OpenGL 4.6, and Vulkan 1.4. The Rubin GPU lists N/A for DirectX, OpenGL, and Vulkan, indicating it is not designed for conventional graphics APIs. The bus interface also differs: PCIe 4.0 x8 for the mobile part versus PCIe 6.0 x16 for the server part.

Power and physical configuration are drastically different. The RTX 4070 Mobile has a TDP of 115 W, uses no external power connectors, and is an IGP with portable device dependent outputs. The Rubin GPU has a TDP of 2300 W, comes as an SXM module, and has no display outputs. The suggested PSU for the Rubin system is 2700 W.

Release timing shows the generational gap. The RTX 4070 Mobile was released in early 2023, while the Rubin GPU is dated to late 2025. The mobile part's predecessor is GeForce 30 Mobile and its successor is GeForce 50 Mobile. The Rubin GPU's predecessor is Server Blackwell with no successor listed.

Head-to-Head Benchmarks

The database contains no direct head-to-head benchmark comparisons between the RTX 4070 Mobile and the Rubin GPU. The Rubin GPU has an empty benchmark array and no nearest rivals, meaning no measured performance data exists for it. The RTX 4070 Mobile, however, has a full set of nine benchmark scores.

In Geekbench OpenCL, the RTX 4070 Mobile scores 109197. In Geekbench Vulkan, it scores 108367. Passmark results show a DirectX 10 score of 116, DirectX 11 score of 179, DirectX 12 score of 85, and DirectX 9 score of 223. The 2D graphics score is 763, the 3D graphics score is 19587, and the GPU compute score is 8399.

The RTX 4070 Mobile's average benchmark score of 27435 places it nearly exactly at the level of the AMD Radeon RX 6700 XT, which scores 27425 with a 0% delta. It trails the NVIDIA GeForce RTX 3090 by half a percent, as that card scores 27565. It leads the NVIDIA RTX PRO 4000 Blackwell by 1.1%, which scores 27135. The AMD Radeon Pro Vega 20 scores 27839, putting it 1.5% ahead of the mobile part.

The Rubin GPU's lack of any benchmark data means no comparative wins can be established. The wins count shows 0 for the RTX 4070 Mobile and 0 for the Rubin GPU, reflecting the absence of head-to-head tests. The 50th percentile ranking for the Rubin GPU is a default placement rather than a measured result, since its average benchmark score is 0.

Without measured scores, the only quantitative comparisons come from specification-derived rates. The RTX 4070 Mobile delivers 15.62 TFLOPS FP32 and the same 15.62 TFLOPS FP16 on a 1:1 ratio. The Rubin GPU delivers 130.0 TFLOPS FP32 and 260.0 TFLOPS FP16 on a 2:1 ratio. Texture rate is 244.1 GTexel/s for the mobile part versus 2,031.2 GTexel/s for the Rubin part. Pixel rate is 81.36 GPixel/s for the mobile part versus 54.41 GPixel/s for the Rubin part.

The memory bandwidth gap is enormous: 256.0 GB/s for the RTX 4070 Mobile versus 22.1 TB/s for the Rubin GPU. The bus width difference from 128 bit to 16384 bit explains this. These figures indicate the Rubin GPU is built for massive data throughput rather than interactive graphics.

Specification Differences

| Specification | RTX 4070 Mobile | Rubin GPU |

|---|---|---|

| Chip | AD106 | GR100 |

| Architecture | Ada Lovelace | Rubin |

| Generation | GeForce 40 Mobile | Server Rubin (Rxx) |

| Process node | 5 nm | 3 nm |

| Transistors | 22,900 million | 336,000 million |

| Die size | 188 mm² | 1456 mm² |

| Transistor density | 121.8M / mm² | 230.8M / mm² |

| Base clock | 1395 MHz | 700 MHz |

| Boost clock | 1695 MHz | 2267 MHz |

| Memory clock | 2000 MHz, 16 Gbps effective | 2695 MHz, 10.8 Gbps effective |

| Memory size | 8 GB | 288 GB |

| Memory type | GDDR6 | HBM4 |

| Memory bus | 128 bit | 16384 bit |

| Memory bandwidth | 256.0 GB/s | 22.1 TB/s |

| Shading units | 4608 | 28672 |

| TMUs | 144 | 896 |

| ROPs | 48 | 24 |

| RT cores | 36 | N/A |

| Tensor cores | 144 | 896 |

| Pixel rate | 81.36 GPixel/s | 54.41 GPixel/s |

| Texture rate | 244.1 GTexel/s | 2,031.2 GTexel/s |

| FP32 | 15.62 TFLOPS | 130.0 TFLOPS |

| FP16 | 15.62 TFLOPS (1:1) | 260.0 TFLOPS (2:1) |

| TDP | 115 W | 2300 W |

| Slot width | IGP | SXM Module |

| Power connectors | None | Not listed |

| Suggested PSU | Not listed | 2700 W |

| Bus interface | PCIe 4.0 x8 | PCIe 6.0 x16 |

| Display outputs | Portable Device Dependent | No outputs |

| DirectX | 12 Ultimate (12_2) | N/A |

| OpenGL | 4.6 | N/A |

| Vulkan | 1.4 | N/A |

| Release date | 2023-01-02 | 2025-12-31 |

| Predecessor | GeForce 30 Mobile | Server Blackwell |

| Successor | GeForce 50 Mobile | Not listed |

| Production status | Active | Active |

The Rubin GPU's RT core count is listed as null in the database, meaning no figure is recorded. The power connector field is also null. The suggested PSU of 2700 W applies to the overall system configuration, not just the GPU module.

Where Each One Wins

The RTX 4070 Mobile wins in any scenario requiring conventional graphics rendering. It supports DirectX 12 Ultimate, OpenGL 4.6, and Vulkan 1.4, while the Rubin GPU supports none of these APIs. The mobile part has 48 ROPs versus 24 on the Rubin GPU, giving it a higher pixel rate of 81.36 GPixel/s versus 54.41 GPixel/s. Its 36 ray tracing cores provide hardware RT support that the Rubin GPU does not list.

The RTX 4070 Mobile also wins on power efficiency for its class. Its 115 W TDP is a fraction of the Rubin GPU's 2300 W. The mobile part requires no external power connectors and works as an IGP, making it suitable for portable devices. The Rubin GPU needs a 2700 W suggested PSU and comes as an SXM module, which is a server form factor.

The database shows the RTX 4070 Mobile performing at the level of desktop GPUs like the RX 6700 XT and RTX 3090, with deltas within 1.5%. Its 73rd percentile ranking indicates solid placement among all GPUs. The 19587 Passmark G3D score and 8399 compute score provide concrete data points for mobile gaming and general workloads.

The Rubin GPU wins decisively on raw compute throughput. Its FP32 performance of 130.0 TFLOPS is over 8 times the mobile part's 15.62 TFLOPS. FP16 performance of 260.0 TFLOPS is over 16 times the mobile part's 15.62 TFLOPS. The texture rate of 2,031.2 GTexel/s is over 8 times higher. Memory bandwidth of 22.1 TB/s is over 86 times higher.

The Rubin GPU wins on memory capacity and type. Its 288 GB of HBM4 dwarfs the 8 GB of GDDR6 on the mobile part. The 16384-bit bus is 128 times wider. These specifications suit large-scale compute workloads, AI training, and data center applications where massive datasets must reside close to the processor.

The Rubin GPU also wins on process technology and transistor scale. The 3 nm node and 336,000 million transistors represent the leading edge, with a density of 230.8M per mm². Its boost clock of 2267 MHz exceeds the mobile part's 1695 MHz despite the lower base clock. The PCIe 6.0 x16 interface provides more host bandwidth than PCIe 4.0 x8.

The tensor core count of 896 versus 144 indicates the Rubin GPU is oriented toward matrix operations. The 2:1 FP16 ratio suggests mixed-precision workloads. The absence of display outputs confirms it is not intended for graphics output. The N/A API entries reinforce this: it is a compute accelerator, not a rendering device.

The release dates show the Rubin GPU as a later product, succeeding Server Blackwell. The RTX 4070 Mobile belongs to the GeForce 40 Mobile line with a clear successor in GeForce 50 Mobile. Both are listed as active in production status, but they serve separate markets with no overlap in use cases.

DETAILED SPECIFICATIONS

SPECIFICATION
RTX 4070 Mobile
Rubin GPU
Core Specs
Shading Units
4,608
28,672 +522.2%
Shaders
4,608
28,672 +522.2%
TMUs
144
896 +522.2%
ROPs
48
24 -50.0%
SM Count
36
224 +522.2%
Clocks
Base Clock
1395 MHz
700 MHz
Boost Clock
1695 MHz
2267 MHz
Memory Clock
2000 MHz 16 Gbps effective
2695 MHz 10.8 Gbps effective
Memory
Memory Size
8 GB
288 GB
VRAM (MB)
8,192
294,912 +3500.0%
Memory Type
GDDR6
HBM4
Memory Bus
128 bit
16384 bit
Bandwidth
256.0 GB/s
22.1 TB/s
Cache
L1 Cache
128 KB (per SM)
256 KB (per SM)
L2 Cache
32 MB
128 MB
Performance
Pixel Rate
81.36 GPixel/s
54.41 GPixel/s
Texture Rate
244.1 GTexel/s
2,031.2 GTexel/s
FP32 (TFLOPS)
15.62 TFLOPS
130.0 TFLOPS
FP64 (TFLOPS)
244.1 GFLOPS (1:64)
32.50 TFLOPS (1:4)
FP16 (TFLOPS)
15.62 TFLOPS (1:1)
260.0 TFLOPS (2:1)
AI/RT
RT Cores
36
Tensor Cores
144
896 +522.2%
Power
TDP
115 W
2300 W
TDP (W)
115
2,300 +1900.0%
Suggested PSU
2700 W
Power Connectors
None
Architecture
Architecture
Ada Lovelace
Rubin
GPU Name
AD106
GR100
Generation
GeForce 40 Mobile
Server Rubin (Rxx)
Process Size
5 nm
3 nm
Transistors
22,900 million
336,000 million
Die Size
188 mm²
1456 mm²
Foundry
TSMC
TSMC
Density
121.8M / mm²
230.8M / mm²
API Support
DirectX
12 Ultimate (12_2)
OpenGL
4.6
Vulkan
1.4
OpenCL
3.0
3.0
CUDA
8.9
10.7
Shader Model
6.8
Physical
Slot Width
IGP
SXM Module
Outputs
Portable Device Dependent
No outputs
Bus Interface
PCIe 4.0 x8
PCIe 6.0 x16
Other
Production
Active
Active
Predecessor
GeForce 30 Mobile
Server Blackwell
Successor
GeForce 50 Mobile
View GeForce RTX 4070 Mobile Details View Rubin GPU Details