GPU Wiki

GeForce RTX 4070 SUPER vs GeForce RTX 4070 Ti SUPER

Swap Clear comparison
Edit GPUs
Add a third or fourth GPU

Key Differences

  • No game was measured on every one of these cards on one shared test system, so there is no speed comparison to make.
  • RTX 4070 Ti SUPER has the most memory: 16 GB against 12 GB.
  • RTX 4070 Ti SUPER has the most memory bandwidth: 672.3 GB/s against 504.2 GB/s.
  • RTX 4070 SUPER draws the least power: 220 W against 285 W.
Spec GeForce RTX 4070 SUPER GeForce RTX 4070 Ti SUPER
Overview
Product name NVIDIA GeForce RTX 4070 SUPER NVIDIA GeForce RTX 4070 Ti SUPER
Architecture Ada Lovelace Ada Lovelace
Generation / family GeForce RTX 40 Series GeForce RTX 40 Series
GPU codename AD104 AD103
Launch date 2024-01-17 2024-01-24
Market segment Desktop in GeForce RTX 40 Series Desktop, High End in GeForce RTX 40 Series
Board type / entry type Reference GPU Reference GPU
GPU
Process node TSMC 5 nm TSMC 5 nm
Die size 294 mm² 379 mm²
Transistor count 35.8 billion 45.9 billion
GPU chip AD104 AD103
SM / multiprocessor count 56 66
CUDA Cores 7168 8448
Tensor Cores 224 264
RT Cores 56 66
ROPs 80 96
TMUs 224 264
Clocks
Base / Core clock 1,980 MHz 2,340 MHz
Boost clock 2,475 MHz 2,610 MHz
Memory clock 1,313 MHz 1,313 MHz
Memory
Memory size 12 GB 16 GB
Memory type GDDR6X GDDR6X
Memory bus width 192-bit 256-bit
Memory bandwidth 504.2 GB/s 672.3 GB/s
Bus / Interface
Bus interface PCIe 4.0 x16 PCIe 4.0 x16
AGP / PCI / PCIe distinction PCIe PCIe
Power
TDP / TBP / board power 220 W 285 W
Recommended PSU 650 W 700 W
Power connectors 1x PCIe 8-pin or 1x PCIe Gen 5 16-pin 1x PCIe Gen 5 16-pin
Slot width Dual-slot, 267 mm length, 42 mm width Triple-slot, 310 mm length, 61 mm width
Display / Media
Display outputs 1x HDMI 2.1, 3x DisplayPort 1.4a 1x HDMI 2.1, 3x DisplayPort 1.4a
Features
DirectX support DirectX 12.2 DirectX 12.2
OpenGL support OpenGL 4.6 OpenGL 4.6
Vulkan support Vulkan 1.4 Vulkan 1.4
CUDA support 8.9 8.9
Shader Model 6.8 6.8
Ray tracing support Hardware RT units present Hardware RT units present
Performance / Benchmarks
FP32 throughput 35.48 TFLOPS 44.1 TFLOPS
FP16 throughput 35.48 TFLOPS 44.1 TFLOPS
Muted cells indicate unavailable, not applicable, or not directly comparable specifications.

Frame rates compare runs of the same game at the same quality preset on one test system. DLAA counts as native resolution, because it applies anti-aliasing at full render resolution rather than upscaling; DLSS and FSR upscaling are never counted as native. Where a capture named ray tracing without a readable on or off value, it is read as on.