NVIDIA A100 80 GB

Train and serve models with twice the memory of the A100 40 GB. Choose it when your workload needs more room for weights, batches or inference cache.

From $1.49/GPU/hour · USD on-demand rate · Storage extra

Workloads

What to run on A100 80 GB.

Choose for your workload’s memory needs, software support and measured runtime.

Larger training batches

Fit more training state on each GPU than the 40 GB variant. Actual batch size depends on your model and precision.

Model serving

Leave room for KV cache and concurrent requests as well as model weights.

Existing CUDA workloads

Run established Ampere-compatible training and inference software in a template or your own VM environment.

Choose your setup

Start with the environment you need.

Use a preconfigured container or manage your operating system in a GPU VM. The dashboard shows current GPU and region availability.

GPU Templates

Start with PyTorch, ComfyUI or another supported environment. Drivers and framework dependencies are preconfigured.

Explore templates

GPU VMs

Get SSH and full root access when you need control over the operating system and runtime. Available configurations vary by GPU and region.

Explore VMs

Persistent storage

Keep datasets and checkpoints on a filesystem you can attach to your own instances in the same region.

Explore filesystems

Sizing

Start with memory. Then measure performance.

Model weights are only part of the working set. Leave room for everything the job needs while it runs.

For inference

Account for model precision, context length, KV cache and concurrent requests. A model loading successfully does not tell you how much traffic it can serve.

For training

Include activations, gradients and optimizer state. Batch size, sequence length and checkpointing change memory use.

For multiple GPUs

GPU memory is not automatically pooled. Use a framework and parallelism strategy that distribute the workload across devices.

Hardware specifications: NVIDIA A100 80 GB. Software and workload affect realized performance.

Compare options

Compare memory and hourly rates.

Use this as a shortlist, then test your workload. Lower hourly pricing does not always mean a lower total job cost.

Published USD on-demand rates per GPU; storage extra. Availability and regional prices vary.
GPUMemoryArchitectureFrom / GPU / hour
NVIDIA H200141 GB HBM3eHopper$3.99
NVIDIA H10080 GB HBM3Hopper$2.69
NVIDIA B200180 GB per GPU¹BlackwellRequest a quote
NVIDIA RTX PRO 600096 GB GDDR7Blackwell$1.89
NVIDIA A100 80 GB80 GB HBM2eAmpere$1.49
NVIDIA A100 40 GB40 GB HBM2Ampere$0.89
NVIDIA L424 GB GDDR6Ada Lovelace$0.44

Before you start

Common questions.

Practical details for choosing and using this product.

How do I get started with NVIDIA A100 80 GB?

Choose a template or VM, select the GPU and region, and review the configuration and price in the dashboard before launch.

Is storage included in the GPU rate?

Storage is billed separately. Retained storage continues to incur charges while an instance is paused. Review the full configuration price before launching.

Will my model fit on one GPU?

This GPU has 80 GB HBM2e. Fit depends on weights, precision, framework overhead and workload state. For inference, also account for context and concurrency; for training, include activations and optimizer state.

Can I get more than eight GPUs?

For workloads spanning nodes, explore GPU clusters. Reserved capacity can be planned at 128, 256, 1,024 GPUs and beyond, subject to configuration and availability.

Get started with NVIDIA A100 80 GB.

Launch a GPU instance or talk to our team about the right setup for your workload.