NVIDIA A100 Specs & Datasheet (80GB SXM · HGX & DGX A100)
TL;DR
The NVIDIA A100 SXM 80GB is an Ampere-generation data-center GPU: 80 GB of HBM2e at 2.04 TB/s per GPU, about 312 TFLOPS of dense FP16 Tensor compute, 400 W, connected over third-generation NVLink at 600 GB/s across an 8-GPU domain. It ships in two 8-GPU forms — the HGX A100 baseboard (integrator-built) and the NVIDIA DGX A100 640GB system, which is exactly eight of these GPUs (640 GB HBM2e per node). A100 is a prior-generation part (Hopper H100/H200 and Blackwell are the current silicon); it is still widely deployed and available, with configuration and pricing confirmed at quote.
On this page
What the A100 is
The NVIDIA A100 is a data-center GPU built on NVIDIA’s Ampere architecture. The mainstream data-center variant is the A100 SXM 80GB — an SXM4-form GPU with 80 GB of on-package HBM2e memory, used for AI training, inference, and HPC.
The A100 is almost always deployed as an 8-GPU building block, and it appears in two forms:
- HGX A100 — the 8-GPU NVLink baseboard that integrators (e.g. Supermicro) drop into their own server chassis. You buy it as part of an OEM system.
- NVIDIA DGX A100 — NVIDIA’s own turnkey appliance. The DGX A100 640GB is exactly eight A100 SXM 80GB GPUs (640 GB HBM2e per system) on the same 8-GPU NVLink domain — the DGX is the integrated system, the A100 SXM is the chip inside it.
A100 is a prior-generation part: it predates the Hopper H100/H200 and the current Blackwell generation. It remains widely deployed and available on the secondary market as fleets rotate to newer silicon — a proven, brand-recognized platform rather than the latest one.
Full spec sheet
The NVIDIA A100 SXM 80GB (Ampere) specification. The per-GPU figures below are fixed characteristics of the chip; the 8-GPU column is the same GPU on an 8-GPU NVLink domain — the form the HGX A100 node and the DGX A100 640GB system both take. Chassis, host CPU, cooling, site power, and final configuration vary by integrator and are confirmed at quote.
| Spec | A100 SXM (per GPU) | 8-GPU (HGX / DGX A100) |
|---|---|---|
| GPU architecture | NVIDIA A100 · Ampere | 8× A100 SXM 80GB |
| Form factor | SXM4 | 8-GPU NVLink domain |
| HBM2e memory | 80 GB | 640 GB |
| Memory bandwidth | 2.04 TB/s per GPU | 2.04 TB/s per GPU |
| FP16 Tensor (dense) | ~312 TFLOPS (~0.31 PFLOPS) | ~2.5 PFLOPS (derived) |
| NVLink | 3rd-gen · 600 GB/s per GPU | 8-GPU NVLink domain |
| TDP | 400 W per GPU | Per-system power confirmed at quote |
| Forms | HGX A100 baseboard · DGX A100 | Supermicro HGX · NVIDIA DGX A100 |
Memory & bandwidth (VRAM)
Each A100 SXM carries 80 GB of HBM2e — the on-package GPU memory (VRAM) available to the model — at 2.04 TB/s per GPU of memory bandwidth.
Across the 8-GPU NVLink domain that both the HGX A100 node and the DGX A100 640GB system use, that is 640 GB of HBM2e per node (8 × 80 GB), pooled and addressable over NVLink so a large model and its KV cache can span all eight GPUs. This is why the DGX A100 640GB carries that number in its name: it is eight A100 SXM 80GB GPUs, nothing more exotic.
NVLink
Inside an 8-GPU node the A100s are wired over third-generation NVLink at 600 GB/s per GPU, giving all-to-all connectivity across the 8-GPU NVLink domain so the node behaves as one tightly-coupled accelerator rather than eight discrete cards.
That 8-GPU NVLink domain is the fixed characteristic shared by the HGX A100 baseboard and the DGX A100 system; scale-out networking (InfiniBand adapters, GPU:NIC ratio) is where the integrator builds differ, and is confirmed per build at quote.
Available integrators
The A100 SXM 80GB ships in two 8-GPU forms below — the Supermicro HGX A100 8-GPU node and the NVIDIA DGX A100 640GB turnkey appliance. Both carry the same A100 SXM 80GB GPU, the same 640 GB of HBM2e per node, and the same 8-GPU third-generation NVLink domain; chassis, host CPU, networking, warranty, and lead time differ by build. Each links to its full specification.
Supermicro
Supermicro AS-4124GO-NART
NVIDIA HGX A100 8-GPU server — 640 GB HBM2e per node on the Supermicro AS-4124GO-NART, the deep secondary-market workhorse for training and inference.
NVIDIA
640 GB
NVIDIA DGX A100 — the 8× A100 SXM4 turnkey appliance with 640 GB HBM2e, a brand-recognized building block for AI training now trading on the secondary market.
A100 vs the current generation
The A100 is Ampere — a prior-generation part. If you want the current silicon, the honest pointer is up a generation:
- Hopper H100 / H200 — the A100’s direct successors. The H200 in particular lifts per-GPU memory to 141 GB of HBM3e at far higher bandwidth, a large gain for memory-bound inference over the A100’s 80 GB HBM2e.
- Blackwell (B200 and up) — the current top of the stack; see Blackwell vs Hopper for how the generations line up.
What keeps the A100 in service is not raw spec leadership but availability and value: it is a proven, brand-recognized 8-GPU building block that is widely deployed and trades actively on the secondary market as fleets rotate to Hopper and Blackwell. For workloads that fit 80 GB per GPU (or 640 GB across a node) it remains a capable training and inference platform — just not the newest one.
Procurement
Both A100 forms — the HGX A100 8-GPU node and the DGX A100 640GB system — are quoted per configuration on request; the figure depends on the build, host CPU, networking, unit grade, and quantity, all confirmed at quote. Tell us the build you are standing up and we will return availability and a facility-fit review. Browse the GPU catalog to compare integrations, or step up a generation to the HGX H200 if you want current Hopper silicon.
Frequently asked questions
What are the DGX A100 specs?
The NVIDIA DGX A100 640GB is a turnkey 8-GPU appliance built from eight A100 SXM 80GB (Ampere) GPUs — 640 GB of HBM2e per system, 2.04 TB/s of memory bandwidth per GPU, about 312 TFLOPS of dense FP16 Tensor compute per GPU (~2.5 PFLOPS across the eight), all on a third-generation NVLink domain at 600 GB/s per GPU. Host CPU, networking, and per-system power are confirmed at quote.
Where can I find the DGX A100 datasheet?
This page is the datasheet: the DGX A100 640GB is exactly 8× the NVIDIA A100 SXM 80GB — 80 GB HBM2e and 2.04 TB/s per GPU, ~312 TFLOPS dense FP16 Tensor, 400 W, third-generation NVLink at 600 GB/s across an 8-GPU domain, for 640 GB of HBM2e per system. The full spec sheet and the DGX A100 vs HGX A100 forms are covered above.
How much memory does the A100 have?
The A100 SXM 80GB has 80 GB of HBM2e per GPU at 2.04 TB/s of memory bandwidth. On an 8-GPU node — the HGX A100 baseboard or the DGX A100 640GB system — that is 640 GB of HBM2e, pooled over NVLink.
Is the A100 a current-generation GPU?
No. The A100 is NVIDIA’s Ampere generation — a prior generation that predates the Hopper H100/H200 and the current Blackwell parts. It is still widely deployed and available as fleets rotate to newer silicon; if you specifically want current silicon, look at the H100/H200 or Blackwell.
How much does an A100 or DGX A100 cost?
Both the HGX A100 node and the DGX A100 640GB system are quoted per configuration on request — the figure depends on the build, host CPU, networking, unit grade, and quantity. Tell us the build you need and we will return availability and a facility-fit review.
Related
Last updated