NVIDIA RTX Pro 6000 Blackwell

96 GB of VRAM on a single card — run 70B-class models, fine-tune with QLoRA, and generate video without offloading. Launch in under a minute.

$1.89/hr on-demand$0.99/hr spotbilled per minute
Key specs
VRAM
96 GB GDDR7 ECC
Memory bandwidth
1,597 GB/s
CUDA cores
24,064
Tensor Cores
5th gen · FP4
Interface
PCIe Gen 5 x16
Configurations
1 · 2 · 4 · 8 GPU
Trusted worldwide

Powering teams that push boundaries

27,000+AI developers
50M+GPU hours served
99.9%Uptime SLA
<90sInstance launch

Trusted by companies including: Tesla, Hugging Face, Kaggle, Zoho, Weights & Biases, upGrad, Saama, Lossfunk

What fits in 96 GB

One card. 70B-class models.

96 GB of GDDR7 is the largest VRAM of any single PCIe card on the platform — here is what that buys you in practice.

Inference

70B-class chat models1 GPU

Llama 3.3 70B, Qwen2.5 72B in FP8

gpt-oss-120b1 GPU

Native MXFP4 weights (~63 GB) with room for KV cache

32B models, full precision1 GPU

Qwen3 32B, DeepSeek-R1-Distill 32B in BF16

Fine-tuning

QLoRA up to 70B1 GPU

4-bit base weights + adapters, single card

LoRA up to 34B in BF161 GPU

CodeLlama 34B, Qwen 32B with long contexts

Full fine-tune 7–8B1 GPU

Llama 3.1 8B with 8-bit optimizers

Image & video generation

FLUX.1, SDXL pipelines1 GPU

Headroom for ControlNets, LoRAs, and large batches

Video models in ComfyUI1 GPU

Wan 2.2 14B and similar — no CPU offload needed

Multi-GPU scale-upup to 8 GPUs

8 × 96 GB = 768 GB VRAM for bigger models and parallel jobs

Works out of the box with vLLM, SGLang, PyTorch, Axolotl, and ComfyUI.

Pricing

Pay for minutes, not months

On-demand and spot are billed by the minute — pause and you pay only for storage. Prices shown are US-region rates.

On-demand
$1.89/hr

Billed per minute. Start and stop any time; pause and pay only for storage.

Spot
$0.99/hr

Interruptible capacity at roughly half the on-demand rate, billed per minute.

1-month plan
$1.49/hr

Reserved capacity, billed for the 30-day term rather than by the minute.

3–12 month plans
Talk to sales

Longer commitments are quoted by sales — pricing varies with duration and cluster size.

Where it sits

Honest positioning against the rest of the fleet

The RTX Pro 6000 has the most VRAM of any single PCIe card here; Hopper cards win on interconnect and HBM bandwidth. Pick the card that matches the job.

NVIDIA RTX Pro 6000 Blackwell$1.89/hr
96 GB GDDR7 · 1.6 TB/s

The single-card sweet spot for 70B-class inference, fine-tuning, and video generation.

80 GB HBM3 · 3.35 TB/s

NVLink and HBM bandwidth for serious multi-GPU training.

141 GB HBM3e · 4.8 TB/s

Largest memory on the platform for frontier-scale models.

80 GB HBM2e · 2.0 TB/s

Proven Ampere workhorse at a lower price point.

How you run it

Managed container or full VM

Every RTX Pro 6000 configuration — 1, 2, 4, or 8 GPUs — is available both ways.

Templates

Managed containers, ready in under a minute

  • PyTorch template with CUDA, drivers, and Jupyter pre-configured
  • SSH and JupyterLab access out of the box
  • Pause and resume — pay only for storage while paused

On-Demand VMs

Full virtual machines with root access

  • Bring your own stack — install anything, run any framework
  • Same per-minute billing and 1–8 GPU configurations
  • Suited for custom drivers, Docker-in-Docker, and long-running services
Full specs

Technical specifications

Verified specifications for the NVIDIA RTX Pro 6000 Blackwell as deployed on JarvisLabs.

ArchitectureNVIDIA Blackwell

Professional (RTX Pro) series, released March 2025

EditionServer Edition

Passively cooled dual-slot server card, as deployed on JarvisLabs

CUDA cores24,064

GB202-based

Tensor Cores752 (5th gen)

FP4 / FP8 / FP16 / BF16 / INT8

RT Cores188 (4th gen)

Ray tracing and 3D rendering

AI performanceUp to 4,000 AI TOPS

FP4 with sparsity

VRAM96 GB GDDR7 (ECC)

Largest single-card VRAM in the RTX line

Memory interface512-bit
Memory bandwidth1,597 GB/s

Server Edition

InterconnectPCIe Gen 5 x16

No NVLink — multi-GPU scales over PCIe

ECC memoryYes

Data integrity for long training runs

Multi-GPU on JarvisLabs1, 2, 4, or 8 GPUs

Templates and On-Demand VMs

27,000+
AI developers
50M+
GPU hours served
99.9%
Uptime SLA
FAQ

Frequently asked questions

Everything you need to know about the NVIDIA RTX Pro 6000 Blackwell on JarvisLabs.

$1.89/hr on-demand and $0.99/hr on spot in US regions, billed per minute — you only pay while an instance is running. A 1-month reserved plan is $1.49/hr, billed for the term; longer commitments are quoted by sales.

Get started

96 GB of Blackwell, one minute away

Launch an RTX Pro 6000 instance from $1.89/hr with per-minute billing. Pause any time and pay only for storage.