NVIDIA GeForce RTX 4070 vs NVIDIA GeForce RTX 4090 Mobile Comparison

NVIDIA
GEFORCE

NVIDIA GeForce RTX 4070

CORE STATE AD104
VRAM 12 GB
CLOCK SPEED 2475 MHz
TDP 200 W
BUS WIDTH 192 bit
ARCHITECTURE Ada Lovelace
nm
PROCESS 5 nm
LAUNCH DATE 2023
VS
NVIDIA
GEFORCE

GeForce RTX 4090 Mobile

CORE STATE AD103
VRAM 16 GB
CLOCK SPEED 1695 MHz
TDP 120 W
BUS WIDTH 256 bit
ARCHITECTURE Ada Lovelace
nm
PROCESS 5 nm
LAUNCH DATE 2023

PERFORMANCE BENCHMARKS

3dmark_3dmark_steel_nomad_dx12
3,854
N/A
geekbench_opencl
154,858
180,831
geekbench_vulkan
174,152
170,774
passmark_directx_10
139
173
passmark_directx_11
244
262
passmark_directx_12
103
107
passmark_directx_9
320
310
passmark_g2d
1,164
984
passmark_g3d
26,927
27,212
passmark_gpu_compute
14,720
12,347

Analysis: NVIDIA GeForce RTX 4070 vs NVIDIA GeForce RTX 4090 Mobile

Architecture Differences

The NVIDIA GeForce RTX 4090 Mobile and NVIDIA GeForce RTX 4070 are both built on the Ada Lovelace architecture, using TSMC's 5 nm process, but they are fundamentally different chips with distinct configurations. The RTX 4090 Mobile uses the AD103 die, which contains 45,900 million transistors across a 379 mm² die, while the RTX 4070 uses the AD104 die with 35,800 million transistors on a 294 mm² die. Transistor density is nearly identical at 121.1M per mm² for the mobile part and 121.8M per mm² for the desktop part, confirming that both chips are produced on the same process node.

The compute resources diverge sharply. The RTX 4090 Mobile carries 9,728 shading units, 304 texture mapping units, and 112 ROPs, while the RTX 4070 has 5,888 shading units, 184 TMUs, and 64 ROPs. Ray tracing hardware follows the same pattern: the mobile flagship has 76 RT cores and 304 tensor cores, whereas the desktop RTX 4070 has 46 RT cores and 184 tensor cores. These differences translate directly into raw throughput: the RTX 4090 Mobile reaches 32.98 TFLOPS for both FP32 and FP16, while the RTX 4070 delivers 29.15 TFLOPS in both precisions.

Memory subsystems are also configured differently. The RTX 4090 Mobile uses 16 GB of GDDR6 memory on a 256-bit bus, yielding 576.0 GB/s of bandwidth. The RTX 4070 features 12 GB of GDDR6X memory on a 192-bit bus, producing 504.2 GB/s. Clock speeds tell another part of the story: the RTX 4090 Mobile runs at a 1335 MHz base and 1695 MHz boost, while the RTX 4070 operates at 1920 MHz base and 2475 MHz boost. The desktop card clocks substantially higher, which helps close the gap in some workloads despite having fewer shaders. The memory clocks also differ, with the mobile part at 2250 MHz (18 Gbps effective) and the desktop part at 1313 MHz (21 Gbps effective).

The physical and power profiles are vastly different. The RTX 4090 Mobile is an integrated graphics processor (IGP) with no power connectors and a 120 W TDP, while the RTX 4070 is a dual-slot card measuring 240 mm by 110 mm by 40 mm, requiring a single 16-pin connector and a 550 W suggested PSU. The RTX 4070 also provides fixed display outputs (1x HDMI 2.1 and 3x DisplayPort 1.4a), whereas the mobile part's outputs are portable device dependent. Both support PCIe 4.0 x16, DirectX 12 Ultimate (12_2), OpenGL 4.6, and Vulkan 1.4. The release dates are close, with the RTX 4090 Mobile arriving on 2023-01-02 and the RTX 4070 on 2023-04-11. The production status differs as well: the mobile part is active, while the desktop part is end-of-life.

The Verdict

The data points to a clear split: the RTX 4090 Mobile wins in most GPU-centric workloads, but the RTX 4070 counters in specific areas, especially compute and 2D performance. Across the nine head-to-head benchmarks, the RTX 4090 Mobile takes five wins, while the RTX 4070 takes four. The average benchmark score puts the RTX 4090 Mobile at 43,667, which places it in the 84th percentile among all GPUs, while the RTX 4070 averages 37,648 and sits in the 81st percentile.

The RTX 4090 Mobile's largest victories are decisive. It leads by 24.5% in Passmark DirectX 10, by 16.8% in Geekbench OpenCL, and by 7.4% in Passmark DirectX 11. It also wins, more narrowly, in Passmark DirectX 12 (3.9%) and Passmark G3D (1.1%). The RTX 4070 strikes back with a 16.1% lead in Passmark GPU Compute, a 15.5% advantage in Passmark G2D, and a 1.9% margin in Geekbench Vulkan, plus a 3.1% win in Passmark DirectX 9.

For users prioritizing raw graphics throughput, especially in DirectX 10 and OpenCL workloads, the RTX 4090 Mobile is the stronger choice. Its 32.98 TFLOPS of FP32 performance, combined with 16 GB of memory and 576.0 GB/s of bandwidth, gives it a clear edge in shader-heavy and memory-intensive scenarios. However, the RTX 4070's compute advantage (14,720 vs 12,347 in Passmark GPU Compute) and its higher boost clock (2475 MHz vs 1695 MHz) make it the better option for tasks that respond well to clock speed and compute-focused architectures.

The RTX 4070 also offers the practical benefit of being a desktop card with fixed display outputs and a dual-slot form factor, while the RTX 4090 Mobile is an IGP whose outputs depend on the portable device. The RTX 4070 carries a launch MSRP of 599 USD, noted once here for reference. The RTX 4090 Mobile has no listed launch MSRP.

FAQ

Q: Which GPU has more shading units?

A: The NVIDIA GeForce RTX 4090 Mobile has 9,728 shading units, compared to 5,888 for the RTX 4070, a difference of 3,840 units.

Q: How much memory bandwidth does each card provide?

A: The RTX 4090 Mobile offers 576.0 GB/s over a 256-bit bus with 16 GB GDDR6, while the RTX 4070 provides 504.2 GB/s over a 192-bit bus with 12 GB GDDR6X.

Q: Which GPU wins in Geekbench OpenCL?

A: The RTX 4090 Mobile scores 180,831 versus 154,858 for the RTX 4070, a 16.8% advantage for the mobile part.

Q: What about Geekbench Vulkan performance?

A: The RTX 4070 wins that test, scoring 174,152 against 170,774, a 1.9% margin for the desktop card.

Q: Which GPU has the higher boost clock?

A: The RTX 4070 boosts to 2475 MHz, well above the RTX 4090 Mobile's 1695 MHz boost clock. The desktop card also has a higher base clock at 1920 MHz versus 1335 MHz.

Q: What is the power draw difference?

A: The RTX 4090 Mobile has a 120 W TDP and requires no power connectors, while the RTX 4070 has a 200 W TDP and needs a single 16-pin connector with a suggested 550 W PSU.

Q: In which benchmark does the RTX 4070 have its largest win?

A: The RTX 4070 leads by 16.1% in Passmark GPU Compute, scoring 14,720 versus 12,347 for the RTX 4090 Mobile.

Specification Differences

| Specification | NVIDIA GeForce RTX 4090 Mobile | NVIDIA GeForce RTX 4070 |

|----------------|-------------------------------|------------------------|

| Chip | AD103 | AD104 |

| Transistors | 45,900 million | 35,800 million |

| Die Size | 379 mm² | 294 mm² |

| Transistor Density | 121.1M / mm² | 121.8M / mm² |

| Base Clock | 1335 MHz | 1920 MHz |

| Boost Clock | 1695 MHz | 2475 MHz |

| Memory Clock | 2250 MHz (18 Gbps effective) | 1313 MHz (21 Gbps effective) |

| Memory Size | 16 GB GDDR6 | 12 GB GDDR6X |

| Memory Bus Width | 256 bit | 192 bit |

| Memory Bandwidth | 576.0 GB/s | 504.2 GB/s |

| Shading Units | 9728 | 5888 |

| TMUs | 304 | 184 |

| ROPs | 112 | 64 |

| RT Cores | 76 | 46 |

| Tensor Cores | 304 | 184 |

| Pixel Rate | 189.8 GPixel/s | 158.4 GPixel/s |

| Texture Rate | 515.3 GTexel/s | 455.4 GTexel/s |

| FP32 / FP16 | 32.98 TFLOPS | 29.15 TFLOPS |

| TDP | 120 W | 200 W |

| Slot Width | IGP | Dual-slot |

| Power Connectors | None | 1x 16-pin |

| Suggested PSU | None | 550 W |

| Display Outputs | Portable Device Dependent | 1x HDMI 2.1, 3x DisplayPort 1.4a |

| Dimensions | Not specified | 240 mm x 110 mm x 40 mm |

| Production Status | Active | End-of-life |

| Release Date | 2023-01-02 | 2023-04-11 |

| Predecessor | GeForce 30 Mobile | GeForce 30 |

| Successor | GeForce 50 Mobile | GeForce 50 |

Head-to-Head Benchmarks

The RTX 4090 Mobile dominates in DirectX 10 rendering, where it scores 173 against the RTX 4070's 139, a 24.5% advantage. This is the single largest margin in the entire comparison. OpenCL performance also favors the mobile part substantially: 180,831 versus 154,858, a 16.8% gap. In DirectX 11, the RTX 4090 Mobile leads by 7.4% (262 vs 244), and in DirectX 12 it edges ahead by 3.9% (107 vs 103). The closest win for the RTX 4090 Mobile comes in Passmark G3D, where it scores 27,212 against 26,927, just 1.1% higher.

The RTX 4070's biggest victory is in Passmark GPU Compute, where it scores 14,720 versus 12,347, a 16.1% margin. This result is notable because it reverses the overall compute hierarchy suggested by FP32 TFLOPS: the RTX 4090 Mobile has higher theoretical FP32 throughput (32.98 TFLOPS) but falls behind in this specific compute test. The RTX 4070 also wins Passmark G2D by 15.5% (1,164 vs 984), indicating stronger 2D graphics performance. In Geekbench Vulkan, the desktop card leads 174,152 to 170,774, a 1.9% edge. The smallest win for the RTX 4070 is in Passmark DirectX 9: 320 versus 310, a 3.1% margin.

The overall average benchmark scores confirm the mobile part's higher standing: 43,667 for the RTX 4090 Mobile versus 37,648 for the RTX 4070. The nearest rivals in the database further contextualize these results. The RTX 4090 Mobile sits 0.8% above the NVIDIA Quadro M6000 (43,301), 0.9% above both the RTX 5050 Mobile (43,268) and the Quadro M6000 24 GB (43,262), and 0.9% below the RTX A6000 (44,075). The RTX 4070 is 0.1% above the Tesla P4 (37,628), 0.4% above the Radeon RX Vega 56 (37,507), 1.3% below the RTX 4080 Mobile (38,135), and 1.3% above the Radeon PRO W6400 (37,157).

Where Each One Wins

The RTX 4090 Mobile is the pick for users who need maximum graphics throughput in DirectX 10, DirectX 11, and OpenCL workloads. Its 24.5% DirectX 10 lead and 16.8% OpenCL advantage are the most pronounced differences in the dataset, and its higher pixel rate (189.8 GPixel/s) and texture rate (515.3 GTexel/s) support this profile. The 16 GB memory capacity and 576.0 GB/s bandwidth also make it the better choice for large textures and high-resolution rendering scenarios, where the extra 4 GB and 71.8 GB/s of bandwidth provide headroom.

The RTX 4070 wins in compute-heavy tasks, as evidenced by its 16.1% lead in Passmark GPU Compute. This result suggests that the desktop card's higher clocks (2475 MHz boost) and its specific compute scheduling give it an advantage in general-purpose GPU workloads, despite having fewer shaders and lower theoretical FP32. Its 15.5% win in Passmark G2D indicates superior 2D graphics performance, which matters for desktop productivity and non-3D rendering tasks. The narrow Vulkan win (1.9%) and DirectX 9 win (3.1%) round out a profile that favors lower-level and compute-oriented APIs.

For gaming and typical 3D workloads, the two cards are much closer than their specifications suggest. The RTX 4090 Mobile wins Passmark G3D by just 1.1%, and the DirectX 12 margin is only 3.9%. The RTX 4070's higher clock speeds compensate for its smaller chip, making it competitive in modern APIs. However, the RTX 4090 Mobile's superior memory bandwidth and larger frame buffer give it an edge in memory-bound scenarios, while the RTX 4070's compute advantage makes it the more versatile choice for mixed-use systems that handle both graphics and compute tasks.

In terms of physical integration, the RTX 4090 Mobile is designed for laptops and compact devices, with a 120 W TDP and no external power requirement. The RTX 4070, as a dual-slot desktop card with a 200 W TDP and a 550 W suggested PSU, requires a full desktop chassis. The RTX 4070 also provides standard display outputs, while the mobile part's outputs depend on the host device. The production status difference (active for the mobile part, end-of-life for the desktop part) suggests that the RTX 4090 Mobile remains in current production, while the RTX 4070 has been phased out.

DETAILED SPECIFICATIONS

SPECIFICATION
RTX 4070
RTX 4090 Mobile
Core Specs
Shading Units
5,888
9,728 +65.2%
Shaders
5,888
9,728 +65.2%
TMUs
184
304 +65.2%
ROPs
64
112 +75.0%
SM Count
46
76 +65.2%
Clocks
Base Clock
1920 MHz
1335 MHz
Boost Clock
2475 MHz
1695 MHz
Memory Clock
1313 MHz 21 Gbps effective
2250 MHz 18 Gbps effective
Memory
Memory Size
12 GB
16 GB
VRAM (MB)
12,288
16,384 +33.3%
Memory Type
GDDR6X
GDDR6
Memory Bus
192 bit
256 bit
Bandwidth
504.2 GB/s
576.0 GB/s
Cache
L1 Cache
128 KB (per SM)
128 KB (per SM)
L2 Cache
36 MB
64 MB
Performance
Pixel Rate
158.4 GPixel/s
189.8 GPixel/s
Texture Rate
455.4 GTexel/s
515.3 GTexel/s
FP32 (TFLOPS)
29.15 TFLOPS
32.98 TFLOPS
FP64 (TFLOPS)
455.4 GFLOPS (1:64)
515.3 GFLOPS (1:64)
FP16 (TFLOPS)
29.15 TFLOPS (1:1)
32.98 TFLOPS (1:1)
AI/RT
RT Cores
46
76 +65.2%
Tensor Cores
184
304 +65.2%
Power
TDP
200 W
120 W
TDP (W)
200
120 -40.0%
Suggested PSU
550 W
Power Connectors
1x 16-pin
None
Architecture
Architecture
Ada Lovelace
Ada Lovelace
GPU Name
AD104
AD103
Generation
GeForce 40
GeForce 40 Mobile
Process Size
5 nm
5 nm
Transistors
35,800 million
45,900 million
Die Size
294 mm²
379 mm²
Foundry
TSMC
TSMC
Density
121.8M / mm²
121.1M / mm²
API Support
DirectX
12 Ultimate (12_2)
12 Ultimate (12_2)
OpenGL
4.6
4.6
Vulkan
1.4
1.4
OpenCL
3.0
3.0
CUDA
8.9
8.9
Shader Model
6.8
6.8
Physical
Slot Width
Dual-slot
IGP
Length
240 mm 9.4 inches
Height
110 mm 4.3 inches
Outputs
1x HDMI 2.13x DisplayPort 1.4a
Portable Device Dependent
Bus Interface
PCIe 4.0 x16
PCIe 4.0 x16
Other
Launch Price
599 USD
Production
End-of-life
Active
Predecessor
GeForce 30
GeForce 30 Mobile
Successor
GeForce 50
GeForce 50 Mobile
View GeForce RTX 4070 Details View GeForce RTX 4090 Mobile Details