A100 80 GB Best value
Ampere, HBM2e
- Memory bandwidth
- ~
2.0 TB/s - Precisions
- FP16, BF16, TF32, INT8
- Fits at
4-bit - Up to about 70B
- Best for
Fine-tuning , steady inference

A100
HBM or ECC memory and NVLink for training,
01Plans & pricing
Every GPU is dedicated to you, with fast NVMe for datasets and checkpoints.
| GPU | VRAM | CPU | RAM | Storage | Locations | Price | Order |
|---|---|---|---|---|---|---|---|
|
|
$355.99/mo
Cheapest competitor |
Deploy |
|||||
|
L40S
Inference at scale
48 GB VRAM |
$418.99/mo
Cheapest competitor |
Deploy L40S | |||||
|
|
$581.99/mo
Cheapest competitor |
Deploy |
|||||
|
|
$1,096.99/mo
Cheapest competitor |
Deploy |
|||||
|
|
$2,348.99/mo
Cheapest competitor |
Deploy |
No GPU plan in this jurisdiction yet: or compare them.
Need eight GPUs or InfiniBand between servers? GPU and dedicated customers can ask for a custom build by ticket from the client area.
02Choose
Both have
Ampere, HBM2e
Hopper SXM5, HBM3
The L40S (
03Sizing
Our starting points for the most common training and serving jobs.
| Job | Our pick | Why |
|---|---|---|
| LoRA | L40S or A100 | |
| Full | ||
| Serving 70B at | A100 or H100 | About |
| Serving 70B at | About | |
| 405B at |
Memory figures for the weights are in our VRAM table; context windows and batch size add to them.
04Use cases
Jobs where HBM bandwidth, ECC memory and NVLink are worth the higher monthly price.
LoRA, QLoRA and full
vLLM endpoints for
Research runs on up to four
Large image, video or embedding jobs that run for days.
05Jurisdictions
Each server has its own list; the plans above show exactly where.
Latency figures are estimates from distance, not measurements. Every jurisdiction has the same prices. How we estimate latency
06Offshore, in writing
What we promise is written into our policies, not just our marketing.
Answered with our policy, never enforced. Only a local
IdentityEmail only
No name, address, phone or ID document, ever. A private or disposable address is fine.
Privacy policyPayment5 cryptocurrencies
Bitcoin, Ethereum, Monero, USDT and Solana, paid
TransparencySigned canary
A
07FAQ
Another question? The full FAQ answers what people ask before an order: payments, privacy, complaints and support.
Read the full FAQOur H100 servers use the SXM5 form factor, with HBM3 at
Yes. Every GPU is dedicated to your server. There is no
The nvidia-smi
Between 1 and
GPU servers are not refundable once delivered, because the hardware is set aside for you. Check the specifications and the GPU images guide before you order.