GPU Specifications / NVIDIA

NVIDIA A100 80GB

Ampere datacenter GPU with 80 GB of HBM2e memory, 2039 GB/s of bandwidth and up to 624 TFLOPS of FP16 tensor compute.

Key Specifications

Memory

80 GB HBM2e

Memory Bandwidth

2,039 GB/s

TDP

400 W

Architecture

Ampere

Interconnect

NVLink 3.0 · 600 GB/s

Est. On-demand Price

~$4.50/h

Hourly rates are indicative on-demand estimates; actual pricing varies by provider and commitment.

Compute Performance

PrecisionPeak throughput
FP64 (double precision)9.7 TFLOPS
FP32 (single precision)19.5 TFLOPS
FP2439 TFLOPS
FP16 (tensor)624 TFLOPS
INT8 (tensor)1,248 TFLOPS
INT4 (tensor)2,496 TFLOPS

Tensor figures use the vendor's peak numbers (with structured sparsity where supported).

System Requirements

Recommended CPU

AMD EPYC 7763 or Intel Xeon Platinum 8380

Max VRAM per node (8 GPUs)

640 GB

System RAM (min / recommended)

256 / 512 GB

Minimum PSU

1000 W

LLMs on the A100 80GB

Number of A100 80GB GPUs needed to serve popular models at 8-bit quantization (including 20% overhead for activations and KV cache).

ModelParamsVRAM (8-bit)GPUs needed
GPT-5.6 Sol2400B2682 GB34x A100 80GB
GPT-5 Flagship2100B2347 GB30x A100 80GB
GPT-5.6 Luna1400B1565 GB20x A100 80GB
Kimi K3 (1.2T)1200B1341 GB17x A100 80GB
Kimi K2.6 (1T)1000B1118 GB14x A100 80GB
GPT-5.6 Terra800B894 GB12x A100 80GB
DeepSeek V4 Pro (671B)671B750 GB10x A100 80GB
Llama 4 Behemoth (500B)500B559 GB7x A100 80GB
Claude 5 Fable (480B)480B536 GB7x A100 80GB
GLM 5.2 (400B)400B447 GB6x A100 80GB
Grok 4.5350B391 GB5x A100 80GB
Claude 4.8 Opus (300B)300B335 GB5x A100 80GB
Muse Spark 1.1300B335 GB5x A100 80GB
Grok 4270B302 GB4x A100 80GB
Gemini 3.1 Pro250B279 GB4x A100 80GB
Qwen 3.7 Max (235B)235B263 GB4x A100 80GB
Mistral Large 3 (200B)200B224 GB3x A100 80GB
Grok 3 Mini190B212 GB3x A100 80GB
Claude 5 Sonnet (175B)175B196 GB3x A100 80GB
Gemini 3.5 Flash150B168 GB3x A100 80GB
Gemini 2.5 Flash140B156 GB2x A100 80GB
Llama 4 Maverick (128B)128B143 GB2x A100 80GB
DeepSeek V4 Flash (120B)120B134 GB2x A100 80GB
Qwen 3.6 Plus (110B)110B123 GB2x A100 80GB
Nova Premier (80B)80B89 GB2x A100 80GB
Qwen 3 Coder-Next (80B)80B89 GB2x A100 80GB
Claude 4.5 Haiku (70B)70B78 GB1x A100 80GB
Llama 3.3 Instruct (70B)70B78 GB1x A100 80GB
Mistral Medium 3.5 (70B)70B78 GB1x A100 80GB
Yi 1.5 (40B)40B45 GB1x A100 80GB
Nova Core (34B)34B38 GB1x A100 80GB
DeepSeek V3.1 (32B)32B36 GB1x A100 80GB
Gemma 3 (27B)27B30 GB1x A100 80GB
Mistral Small 4 (24B)24B27 GB1x A100 80GB
Yi 1.5 (15B)15B17 GB1x A100 80GB
Phi 4 (14B)14B16 GB1x A100 80GB
Nova Lite (12B)12B13 GB1x A100 80GB
Llama 3.2 Instruct (11B)11B12 GB1x A100 80GB
Gemma 3 (9B)9B10 GB1x A100 80GB
Yi 1.5 Lite (9B)9B10 GB1x A100 80GB
Phi 4 Mini (7B)7B8 GB1x A100 80GB
Phi 3.5 (3.8B)3.8B4 GB1x A100 80GB

Frequently Asked Questions

How much VRAM does the NVIDIA A100 80GB have?

The NVIDIA A100 80GB has 80 GB of HBM2e memory with 2039 GB/s of memory bandwidth.

Which LLMs can run on a single A100 80GB?

At 8-bit quantization, a single A100 80GB (80 GB) can serve models up to roughly 70B parameters, such as Claude 4.5 Haiku (70B). Larger models require multiple GPUs or more aggressive quantization.

How much does it cost to rent a NVIDIA A100 80GB?

On-demand cloud pricing for the A100 80GB is around $4.50/hour, i.e. about $3,285/month running 24/7. Actual prices vary by provider, region, and commitment.

What are the power and system requirements of the NVIDIA A100 80GB?

The A100 80GB has a TDP of 400W. A power supply of at least 1000W per GPU is recommended. Recommended host CPUs: AMD EPYC 7763 or Intel Xeon Platinum 8380.