Inference & rendering (2022-2024)

NVIDIA
Ada Lovelace

NVIDIA Ada Lovelace GPUs are optimized for AI inference, graphics, and video processing. Featuring GDDR6 memory and PCIe Gen4, they offer excellent price-performance for inference workloads and visual computing.

GPU Models in this Family

Click any card to expand detailed specifications

L40S

L40S

48 GB GDDR6864 GB/sPCIe Gen4

Datacenter GPU optimized for AI inference and graphics. Successor to A40 with 2× performance.

AI inferenceGraphicsRenderingVideo processing
Architecture
Ada Lovelace
VRAM
48 GB GDDR6
Memory Bandwidth
864 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
733 TFLOPS (with sparsity)
FP32 Performance
91 TFLOPS
FP64 Performance
1.42 TFLOPS
TDP
350W
CUDA Cores
18,176
Tensor Cores
568
Process Node
TSMC 5N
Form Factor
PCIe Gen4
L40

L40

48 GB GDDR6864 GB/sPCIe Gen4

Datacenter GPU for visual computing and rendering. Lower AI performance than L40S.

GraphicsRenderingVirtual workstationsVideo processing
Architecture
Ada Lovelace
VRAM
48 GB GDDR6
Memory Bandwidth
864 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
362 TFLOPS (with sparsity)
FP32 Performance
90 TFLOPS
FP64 Performance
1.40 TFLOPS
TDP
300W
CUDA Cores
18,176
Tensor Cores
568
Process Node
TSMC 5N
Form Factor
PCIe Gen4
L4

L4

24 GB GDDR6300 GB/sPCIe Gen4 (single-slot)

Low-power, single-slot GPU for AI inference and video processing. 72W TDP — no extra power connector needed.

AI inferenceVideo processingEdge computingLow-power workloads
Architecture
Ada Lovelace
VRAM
24 GB GDDR6
Memory Bandwidth
300 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
242 TFLOPS (with sparsity)
FP32 Performance
30 TFLOPS
FP64 Performance
0.47 TFLOPS
TDP
72W
CUDA Cores
7,680
Tensor Cores
240
Process Node
TSMC 5N
Form Factor
PCIe Gen4 (single-slot)
L20

L20

48 GB GDDR6864 GB/sPCIe Gen4

China market variant of the L40. 48GB GDDR6 for AI inference and graphics. Meets US export controls while delivering high performance for data center workloads.

AI inferenceGraphicsVideo processing (China market)
Architecture
Ada Lovelace
VRAM
48 GB GDDR6
Memory Bandwidth
864 GB/s
Interconnect
PCIe Gen4
FP16 Performance
119 TFLOPS (with Tensor Cores)
FP32 Performance
59.8 TFLOPS
FP64 Performance
0.93 TFLOPS
TDP
275W
CUDA Cores
14,592
Tensor Cores
456
Process Node
TSMC 4N
Form Factor
PCIe Gen4
L2

L2

24 GB GDDR6300 GB/sPCIe Gen4

China market variant of the L4. 24GB GDDR6 in a low-profile 150W form factor for AI inference and video workloads in compliant regions.

AI inferenceVideo processing (China market)
Architecture
Ada Lovelace
VRAM
24 GB GDDR6
Memory Bandwidth
300 GB/s
Interconnect
PCIe Gen4
FP16 Performance
61 TFLOPS (with Tensor Cores)
FP32 Performance
30.3 TFLOPS
FP64 Performance
0.47 TFLOPS
TDP
150W
CUDA Cores
7,680
Tensor Cores
240
Process Node
TSMC 4N
Form Factor
PCIe Gen4
RTX 6000 Ada

RTX 6000 Ada

48 GB GDDR6960 GB/sPCIe Gen4

Workstation GPU for professional AI, rendering, and graphics. 48GB GDDR6 with ECC support.

Workstation AIRenderingGraphicsVideo processing
Architecture
Ada Lovelace
VRAM
48 GB GDDR6
Memory Bandwidth
960 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
1,466 TFLOPS (with sparsity)
FP32 Performance
91 TFLOPS
FP64 Performance
1.42 TFLOPS
TDP
300W
CUDA Cores
18,176
Tensor Cores
568
Process Node
TSMC 5N
Form Factor
PCIe Gen4
RTX 5000 Ada

RTX 5000 Ada

32 GB GDDR6576 GB/sPCIe Gen4

Mid-range Ada Lovelace workstation GPU. 32GB GDDR6 for professional rendering and AI workloads.

Workstation AIRenderingGraphicsVirtual workstations
Architecture
Ada Lovelace
VRAM
32 GB GDDR6
Memory Bandwidth
576 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
733 TFLOPS (with sparsity)
FP32 Performance
46 TFLOPS
FP64 Performance
0.72 TFLOPS
TDP
250W
CUDA Cores
12,288
Tensor Cores
384
Process Node
TSMC 5N
Form Factor
PCIe Gen4
RTX 4500 Ada

RTX 4500 Ada

24 GB GDDR6432 GB/sPCIe Gen4

Mid-range Ada Lovelace workstation GPU. 24GB GDDR6 for professional rendering and AI workloads.

Workstation AIRenderingGraphics
Architecture
Ada Lovelace
VRAM
24 GB GDDR6
Memory Bandwidth
432 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
550 TFLOPS (with sparsity)
FP32 Performance
34 TFLOPS
FP64 Performance
0.53 TFLOPS
TDP
200W
CUDA Cores
7,680
Tensor Cores
240
Process Node
TSMC 5N
Form Factor
PCIe Gen4
RTX 4000 Ada

RTX 4000 Ada

20 GB GDDR6360 GB/sPCIe Gen4 (single-slot)

Entry-level Ada Lovelace workstation GPU. 20GB GDDR6, single-slot design for compact workstations.

Workstation AIRenderingGraphicsBudget professional workloads
Architecture
Ada Lovelace
VRAM
20 GB GDDR6
Memory Bandwidth
360 GB/s
Interconnect
PCIe Gen4 (64 GB/s)
FP16 Performance
366 TFLOPS (with sparsity)
FP32 Performance
23 TFLOPS
FP64 Performance
0.36 TFLOPS
TDP
130W
CUDA Cores
6,144
Tensor Cores
192
Process Node
TSMC 5N
Form Factor
PCIe Gen4 (single-slot)
RTX 2000 Ada

RTX 2000 Ada

16 GB GDDR6256 GB/sPCIe Gen4

Entry-level Ada Lovelace workstation GPU. 16GB GDDR6 with low 70W TDP for single-slot deployments. Optimized for AI inference, CAD, and content creation workloads.

Entry AI inferenceCADContent creationLight rendering
Architecture
Ada Lovelace
VRAM
16 GB GDDR6
Memory Bandwidth
256 GB/s
Interconnect
PCIe Gen4
FP16 Performance
12 TFLOPS
FP32 Performance
12 TFLOPS
FP64 Performance
0.19 TFLOPS
TDP
70W
CUDA Cores
2,816
Tensor Cores
88
Process Node
TSMC 5N
Form Factor
PCIe Gen4

Quick Comparison

GPU ModelVRAMBandwidthFP16TDPStarting Price
L40S48 GB GDDR6864 GB/s733 TFLOPS (with sparsity)350W
L4048 GB GDDR6864 GB/s362 TFLOPS (with sparsity)300W
L424 GB GDDR6300 GB/s242 TFLOPS (with sparsity)72W
L2048 GB GDDR6864 GB/s119 TFLOPS (with Tensor Cores)275W
L224 GB GDDR6300 GB/s61 TFLOPS (with Tensor Cores)150W
RTX 6000 Ada48 GB GDDR6960 GB/s1,466 TFLOPS (with sparsity)300W
RTX 5000 Ada32 GB GDDR6576 GB/s733 TFLOPS (with sparsity)250W
RTX 4500 Ada24 GB GDDR6432 GB/s550 TFLOPS (with sparsity)200W
RTX 4000 Ada20 GB GDDR6360 GB/s366 TFLOPS (with sparsity)130W
RTX 2000 Ada16 GB GDDR6256 GB/s12 TFLOPS70W

Why Choose Ada Lovelace?

Proven Performance

Ada Lovelace GPUs power the world's largest AI training clusters. Battle-tested in production at scale for LLM training, fine-tuning, and inference.

Cloud-Native Ready

Available across multiple cloud providers with hourly billing. Compare real-time pricing and pick the cheapest option for your workload.

Ecosystem Support

Full compatibility with PyTorch, TensorFlow, JAX, vLLM, TensorRT, and all major ML frameworks. NVLink for multi-GPU scaling.