
A100 SXM
High-performance data center GPU based on Ampere architecture with 80GB HBM2e memory and 6,912 CUDA cores for large-scale AI training and high-performance computing workloads.
GPU Benchmarks
Compare performance across LLMs and image models to find the best GPU for your workload.
Runpod benchmark measurements from September 2026, using vLLM.

High-performance data center GPU based on Ampere architecture with 80GB HBM2e memory and 6,912 CUDA cores for large-scale AI training and high-performance computing workloads.
.avif)
Dual-GPU data center accelerator based on Hopper architecture with 188GB combined HBM3 memory (94GB per GPU) designed specifically for LLM inference and deployment.

High-efficiency LLM processing at 90.98 tok/s.
Runpod benchmark measurements from September 2026, using Hugging Face Diffusers.

Unmatched image gen speed with 49.9 images per minute.

AI image processing at 40.3 images per minute.

Pro-grade performance with 36 images per minute.
Case Studies