🚀 B200 starting at $3.75/hr. The best price you'll find. DC in US West → (Access it from button on top after login).

Get Your B200 →
Start Building
packet.ai/GPUs/NVIDIA B200
In stock · Provisions in ~5 min
NVIDIA Blackwell192 GB HBM3eSXM

NVIDIA B200 GPU

The Blackwell flagship.

The NVIDIA B200 is a flagship Blackwell data-center GPU with 192 GB of HBM3e memory and 8 TB/s of bandwidth, delivering up to 20 petaFLOPS of FP4 AI compute. Available from $3.75/GPU-hour.

from $3.75/GPU-hr· Dynamic $3.75/hr · Dedicated $5.90/hr · Monthly from $2,728/mo
≈ 63% below hyperscaler on-demand
No contractsHourly billingSSH in <5 min
NVIDIA B200 GPU
192GB
HBM3e memory
8TB/s
Memory bandwidth
20PFLOPS
FP4 AI compute (peak)
1.8TB/s
5th-gen NVLink
Specifications

NVIDIA B200 specifications.

SpecificationValueGreat for
GPU architecture
NVIDIA Blackwelldual-die · 208B transistors
One coherent accelerator for trillion-parameter training and inference.
GPU memory
192 GBHBM3e
Fits 70B+ parameter models on a single GPU.
Memory bandwidth
8 TB/s
Feeds Tensor Cores without stalls for LLM inference.
FP4 AI compute
up to 20 PFLOPSpeak, with sparsity
Highest-throughput low-precision inference.
FP8 / FP16 Tensor
~9 / ~2.2 PFLOPS
Accelerates mixed-precision training.
Transformer Engine
2nd generation
Up to 2.5× faster LLM training versus H100.
NVLink
1.8 TB/s5th generation
Fast multi-GPU communication.
Form factor
SXM8-GPU HGX B200 node
Dense data-center nodes.
Availability
On-demandUS & EU regions
Launch in under 5 minutes.
Architecture

Built on NVIDIA Blackwell.

A dual-die design with 208 billion transistors, engineered for trillion-parameter AI and the largest training runs.

5th-gen Tensor Cores + FP4

FP4 and FP8 compute. 2.25 PFLOPS FP4 for the largest transformer models at maximum throughput.

192 GB HBM3e at 8 TB/s

192 GB handles 405B at FP16 natively. 8 TB/s bandwidth, the fastest memory access of any GPU.

NVLink 5.0 (1.8 TB/s)

1.8 TB/s NVLink for tight multi-GPU coupling. Scale to 1,024+ GPUs with full bandwidth coherence.

Transformer Engine 2.0

Native FP4 mixed precision with micro-scaling. 2× the throughput of H100 on transformer inference.

Use Cases

What the B200 is built for.

Large-scale training

Pre-train and fine-tune 70B+ and trillion-parameter models with 2nd-gen Transformer Engine and NVLink scale.

  • 2.5× faster than H100
  • Multi-node NVLink/InfiniBand
  • Trillion-parameter ready

High-throughput inference

Serve the largest LLMs in production with FP4 precision and 192 GB of memory per card.

  • Up to 20 PFLOPS FP4
  • 70B+ models on one GPU
  • Low-latency serving

Fine-tuning

Adapt frontier open models on bursty, hourly capacity without reserving a card.

  • Hourly billing
  • Spin down anytime
  • Full 192 GB VRAM
Detailed Pricing Options

View all pricing tiers and configurations for B200

ConfigurationOn-Demand / hrMonthly3 Months6 MonthsAnnually
Dynamic & DedicatedDedicated only
2× NVIDIA B200Dedicated only

$11.80/hr

$8,614/mo

$11.80/hr eff.

$25,842/3mo

$51,684/6mo

$103,368/yr

4× NVIDIA B200Dedicated only

$23.60/hr

$17,228/mo

$23.60/hr eff.

$51,684/3mo

$103,368/6mo

$206,736/yr

8× NVIDIA B200Dedicated only

$47.20/hr

$34,456/mo

$47.20/hr eff.

$103,368/3mo

$206,736/6mo

$413,472/yr

Deploy hourly →Subscribe to Monthly →
FAQ

NVIDIA B200, answered.

What is the NVIDIA B200?

The B200 is NVIDIA's flagship Blackwell data-center GPU with 192 GB of HBM3e memory at 8 TB/s, delivering up to 20 petaFLOPS of FP4 AI compute.

How much does it cost to rent a B200?

B200 starts at $3.75/GPU-hour Dynamic and $5.90/GPU-hour Dedicated.

How does the B200 compare to the H100?

The B200 is roughly 2.5× faster for LLM workloads than the H100, with 192 GB vs 80 GB memory.

How much memory does the B200 have?

Each B200 has 192 GB of HBM3e at 8 TB/s.

What workloads is the B200 best for?

Large-scale training, fine-tuning frontier models, high-throughput LLM inference with FP4, and HPC.

How fast can I get a B200?

On Dynamic, SSH-ready in under 5 minutes.

Deploy now

Run the Blackwell flagship in under 5 minutes.

Launch a B200 on-demand from $3.75/GPU-hour, $5.90/hr dedicated, or $4,307/mo.

On-demand · hourly billing · US & EU regions

NVIDIA B200from $3.75/GPU-hr
Deploy B200 →