
A100 PCIe
High-performance data center GPU based on Ampere architecture with 80GB HBM2e memory and 6,912 CUDA cores for AI training, machine learning, and high-performance computing workloads.
GPU Benchmarks
Compare performance across LLMs and image models to find the best GPU for your workload.
Runpod benchmark measurements from September 2026, using vLLM.

High-performance data center GPU based on Ampere architecture with 80GB HBM2e memory and 6,912 CUDA cores for AI training, machine learning, and high-performance computing workloads.

Data center GPU based on Ampere architecture with 48GB GDDR6 memory and 10,752 CUDA cores for AI workloads, professional visualization, and virtual workstation applications.

High-efficiency LLM processing at 90.98 tok/s.
Runpod benchmark measurements from September 2026, using Hugging Face Diffusers.

Unmatched image gen speed with 49.9 images per minute.

AI image processing at 40.3 images per minute.

Pro-grade performance with 36 images per minute.
Case Studies