NVIDIAVendor documented
A100 40GB PCIe
Ampere · Ampere · PCIe · 2020
Entry A100 for mainstream servers. Good for mid-size inference and fine-tuning where 40 GB is sufficient.
LLM inferenceFine-tuningVisionHPC
Precision fingerprint
6432t3216BF168i8
Memory
40 GB
HBM2e
Bandwidth
1.555 TB/s
peak
TDP
250 W
air
Max model
~13B
FP16, planning est.
Compute throughput
| FP64 | 19.5 TFLOPS |
| FP32 | 19.5 TFLOPS |
| TF32 | 156 TFLOPS |
| FP16 | 312 TFLOPS |
| BF16 | 312 TFLOPS |
| FP8 | — |
| INT8 | 624 TFLOPS |
Platform & software
InterconnectNVLink bridge — 600 GB/s
PCIePCIe 4.0 x16
Coolingair
MIGSupported
PartitioningUp to 7× MIG
VirtualizationvGPU, MIG
FrameworksCUDA, TensorRT, Triton, vLLM
AvailabilityAWS, Azure, GCP, bare-metal
Known limitations
- ·40 GB limits model size and batch
- ·No FP8