NVIDIA A10 — Ampere Data Center GPU
24 GB
- Architecture
- Ampere
- Generation
- Ampere
- Bandwidth
- 600 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 125 TFLOPS
Get StartedNVIDIA A10G — Ampere Graphics & Inference GPU
22 GB
- Architecture
- Ampere
- Generation
- Ampere
- Bandwidth
- 600 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 70 TFLOPS
Get StartedNVIDIA A16 — Ampere Virtualization GPU
16 GB
Also listed: 2 GB / 4 GB / 8 GB / 16 GB / 32 GB / 64 GB / 128 GB
- Architecture
- Ampere
- Generation
- Ampere
- Bandwidth
- 200 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA A30 — Ampere Tensor Core GPU
24 GB
- Architecture
- Ampere
- Generation
- Ampere
- Bandwidth
- 933 GB/s
- Workload
- Training · Inference
- FP16 dense
- 165 TFLOPS
Get StartedNVIDIA A40 — Ampere Data Center Graphics GPU
48 GB
Also listed: 2 GB / 4 GB / 8 GB / 12 GB / 16 GB / 24 GB / 48 GB
- Architecture
- Ampere
- Generation
- Ampere
- Bandwidth
- 696 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 150 TFLOPS
Get StartedNVIDIA A100 — Ampere Tensor Core AI GPU
80 GB
Also listed: 40 GB / 80 GB
- Architecture
- Ampere
- Generation
- Ampere
- Bandwidth
- 2.0 TB/s
- Workload
- Training · Inference
- FP16 dense
- 312 TFLOPS
Get StartedNVIDIA B200 — Blackwell Tensor Core AI GPU
192 GB
Also listed: 179 GB / 180 GB / 192 GB
- Architecture
- Blackwell
- Generation
- Blackwell
- Bandwidth
- 8 TB/s
- Workload
- Training · Inference
- FP16 dense
- 2,250 TFLOPS
Get StartedNVIDIA B300 — Blackwell Ultra Tensor Core AI GPU
288 GB
Also listed: 268 GB / 270 GB / 288 GB
- Architecture
- Blackwell
- Generation
- Blackwell
- Bandwidth
- 8 TB/s
- Workload
- Training · Inference
- FP16 dense
- 3,500 TFLOPS
Get StartedNVIDIA GB200 — Grace Blackwell AI Superchip
192 GB
- Architecture
- Blackwell
- Generation
- Grace Blackwell
- Bandwidth
- 8 TB/s
- Workload
- Training · Inference
- FP16 dense
- 2,500 TFLOPS
Get StartedNVIDIA GeForce RTX 3070 — Ampere Gaming GPU
8 GB
- Architecture
- Ampere
- Generation
- GeForce 30
- Bandwidth
- 448 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 3080 — Ampere Gaming GPU
10 GB
- Architecture
- Ampere
- Generation
- GeForce 30
- Bandwidth
- 760 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 3080 Ti — Ampere Gaming GPU
12 GB
- Architecture
- Ampere
- Generation
- GeForce 30
- Bandwidth
- 912 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 3090 — Ampere Enthusiast GPU
24 GB
- Architecture
- Ampere
- Generation
- GeForce 30
- Bandwidth
- 936 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 3090 Ti — Ampere Enthusiast GPU
24 GB
- Architecture
- Ampere
- Generation
- GeForce 30
- Bandwidth
- 1.0 TB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 4070 Ti — Ada Lovelace Gaming GPU
12 GB
- Architecture
- Ada Lovelace
- Generation
- GeForce 40
- Bandwidth
- 504 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 4080 — Ada Lovelace Gaming GPU
16 GB
- Architecture
- Ada Lovelace
- Generation
- GeForce 40
- Bandwidth
- 717 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 4080 SUPER — Ada Lovelace Gaming GPU
16 GB
- Architecture
- Ada Lovelace
- Generation
- GeForce 40
- Bandwidth
- 736 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 4090 — Ada Lovelace Enthusiast GPU
24 GB
- Architecture
- Ada Lovelace
- Generation
- GeForce 40
- Bandwidth
- 1.0 TB/s
- Workload
- Graphics
- FP16 dense
- 165 TFLOPS
Get StartedNVIDIA GeForce RTX 5080 — Blackwell Gaming GPU
16 GB
- Architecture
- Blackwell
- Generation
- GeForce 50
- Bandwidth
- 960 GB/s
- Workload
- Graphics
Get StartedNVIDIA GeForce RTX 5090 — Blackwell Enthusiast GPU
32 GB
- Architecture
- Blackwell
- Generation
- GeForce 50
- Bandwidth
- 1.8 TB/s
- Workload
- Graphics
- FP16 dense
- 419 TFLOPS
Get StartedNVIDIA GH200 — Grace Hopper AI Superchip
96 GB
- Architecture
- Hopper
- Generation
- Hopper
- Bandwidth
- 4 TB/s
- Workload
- Training · Inference
- FP16 dense
- 990 TFLOPS
Get StartedNVIDIA H100 — Hopper Tensor Core AI GPU
80 GB
Also listed: 80 GB / 94 GB
- Architecture
- Hopper
- Generation
- Hopper
- Bandwidth
- 3.4 TB/s
- Workload
- Training · Inference
- FP16 dense
- 989 TFLOPS
Get StartedNVIDIA H200 — Hopper Tensor Core AI GPU
141 GB
Also listed: 141 GB / 143 GB
- Architecture
- Hopper
- Generation
- Hopper
- Bandwidth
- 4.8 TB/s
- Workload
- Training · Inference
- FP16 dense
- 989 TFLOPS
Get StartedNVIDIA L4 — Ada Lovelace Inference GPU
24 GB
Also listed: 22 GB / 24 GB
- Architecture
- Ada Lovelace
- Generation
- Ada Lovelace
- Bandwidth
- 300 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 121 TFLOPS
Get StartedNVIDIA L40 — Ada Lovelace Data Center GPU
48 GB
- Architecture
- Ada Lovelace
- Generation
- Ada Lovelace
- Bandwidth
- 864 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 181 TFLOPS
Get StartedNVIDIA L40S — Ada Lovelace AI & Graphics GPU
48 GB
Also listed: 44 GB / 48 GB / 96 GB / 192 GB
- Architecture
- Ada Lovelace
- Generation
- Ada Lovelace
- Bandwidth
- 864 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 181 TFLOPS
Get StartedNVIDIA RTX 2000 Ada — Ada Lovelace Professional GPU
16 GB
- Architecture
- Ada Lovelace
- Generation
- RTX Ada
- Bandwidth
- 224 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX 4000 Ada — Ada Lovelace Professional GPU
20 GB
- Architecture
- Ada Lovelace
- Generation
- RTX Ada
- Bandwidth
- 360 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX 5000 Ada — Ada Lovelace Professional GPU
32 GB
- Architecture
- Ada Lovelace
- Generation
- RTX Ada
- Bandwidth
- 576 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX 6000 Ada — Ada Lovelace Professional GPU
48 GB
- Architecture
- Ada Lovelace
- Generation
- RTX Ada
- Bandwidth
- 960 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 182 TFLOPS
Get StartedNVIDIA RTX 5000 — Turing Professional GPU
32 GB
- Architecture
- Turing
- Generation
- Quadro RTX
- Bandwidth
- 448 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX 6000 — Turing Professional GPU
24 GB
- Architecture
- Turing
- Generation
- Quadro RTX
- Bandwidth
- 672 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX A2000 — Ampere Professional GPU
6 GB
- Architecture
- Ampere
- Generation
- RTX Ampere
- Bandwidth
- 288 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX A4000 — Ampere Professional GPU
16 GB
- Architecture
- Ampere
- Generation
- RTX Ampere
- Bandwidth
- 448 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 77 TFLOPS
Get StartedNVIDIA RTX A4500 — Ampere Professional GPU
20 GB
- Architecture
- Ampere
- Generation
- RTX Ampere
- Bandwidth
- 640 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX A5000 — Ampere Professional GPU
24 GB
- Architecture
- Ampere
- Generation
- RTX Ampere
- Bandwidth
- 768 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 111 TFLOPS
Get StartedNVIDIA RTX A6000 — Ampere Professional GPU
48 GB
- Architecture
- Ampere
- Generation
- RTX Ampere
- Bandwidth
- 768 GB/s
- Workload
- Inference · Graphics
- FP16 dense
- 155 TFLOPS
Get StartedNVIDIA RTX PRO 4000 — Blackwell Professional GPU
24 GB
- Architecture
- Blackwell
- Generation
- RTX PRO Blackwell
- Bandwidth
- 672 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX PRO 4500 — Blackwell Professional GPU
32 GB
- Architecture
- Blackwell
- Generation
- RTX PRO Blackwell
- Bandwidth
- 896 GB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX PRO 5000 — Blackwell Professional GPU
48 GB
- Architecture
- Blackwell
- Generation
- RTX PRO Blackwell
- Bandwidth
- 1.3 TB/s
- Workload
- Inference · Graphics
Get StartedNVIDIA RTX PRO 6000 — Blackwell Professional GPU
96 GB
- Architecture
- Blackwell
- Generation
- RTX PRO Blackwell
- Bandwidth
- 1.8 TB/s
- Workload
- Inference · Graphics
- FP16 dense
- 500 TFLOPS
Get StartedNVIDIA T4 — Turing Tensor Core Inference GPU
16 GB
- Architecture
- Turing
- Generation
- Turing
- Bandwidth
- 320 GB/s
- Workload
- Inference
- FP16 dense
- 65 TFLOPS
Get StartedNVIDIA V100 — Volta Tensor Core AI GPU
32 GB
Also listed: 16 GB / 32 GB
- Architecture
- Volta
- Generation
- Volta
- Bandwidth
- 900 GB/s
- Workload
- Training · Inference
- FP16 dense
- 125 TFLOPS
Get Started