Launch pricing: every plan costs 30% less than the cheapest offshore competitor we track. See the benchmarkEvery plan 30% under the cheapest offshore host

NVIDIA RTX 4090 · 24 GB GDDR6X · Ada Lovelace

Rent an RTX 4090 server.
24 GB for $84.99 a month.

A whole RTX 4090, not a slice of one: 24 GB of GDDR6X at 1,008 GB/s, with 8 vCPU, 64 GB of RAM and 1 TB NVMe, the NVIDIA driver and CUDA preinstalled. That is about $0.116 an hour around the clock, with no meter running. Paid in crypto, no KYC.

  • 24 GB GDDR6X, 1,008 GB/s
  • Ready in 1 to 24 hours
  • No KYC, crypto only
  • Prompts stay on your server

Configure your RTX 4090Ready in 1–24 h

JurisdictionMoldova · Chișinău

PlanCompare plans

Billing cycle

$84.99/mo

Monthly · cancel anytime

Cheapest competitor $121.50−30%

No IDBTC, XMR, USDT +2No setup fee

01Plans & pricing

One RTX 4090 server, 4 jurisdictions.

The card is yours alone, with NVMe storage and an unmetered 1 Gbps port. Quarterly billing saves 5%, yearly 12%.

GPU Servers: 1 plans, monthly prices in USD
GPU VRAM CPU RAM Storage Locations Price Order
  • RTX 4090, 5090, L40S, A100, H100
  • CUDA drivers preinstalled
  • Single and multi-GPU nodes
  • No KYC · BTC, ETH, XMR, USDT, SOL

Need more memory? The RTX 5090 has 32 GB, and two of them 64 GB.

02Specifications

The RTX 4090, in numbers.

NVIDIA’s specifications for the card, and what comes with it in our server.

The RTX 4090, in numbers.
SpecificationRTX 4090 server
ArchitectureNVIDIA Ada Lovelace
CUDA cores16,384
GPU memory24 GB GDDR6X
Memory bandwidth1,008 GB/s
Tensor cores4th generation, with FP8
FP32 compute82.6 TFLOPS
Board power450 W
CPU, RAM, storage8 vCPU · 64 GB · 1 TB NVMe
Network1 Gbps, unmetered
Price$84.99/mo, about $0.116 per hour

03Sizing

What runs on 24 GB.

The highest precision at which a dense language model fits with an 8K-token context and 10% headroom, the figures of our VRAM guide.

What runs on 24 GB.
Model sizeRTX 4090 (24 GB)RTX 5090 (32 GB)
7–8B16-bit16-bit
13–14B8-bit8-bit
32BDoes not fit4-bit
70BDoes not fitDoes not fit
123BDoes not fitDoes not fit

Image models such as SDXL and Flux run comfortably in 24 GB. The full method: how much VRAM does an LLM need?

04Monthly or hourly

When a month beats paying by the hour.

At $84.99 a month, the server costs about $0.116 an hour if it runs non-stop. Against an hourly rental, the month pays for itself after this much use:

When a month beats paying by the hour.
If the hourly rate isOur month equalsBreak-even use
$0.40 per GPU-hour212 h29% of the month
$0.70 per GPU-hour121 h17% of the month
$1.00 per GPU-hour85 h12% of the month

Hourly rates are examples, not quotes from any provider. We bill monthly, quarterly (−5%) or yearly (−12%), from your balance.

05Use cases

What an RTX 4090 is best at.

The value card for models that fit in 24 GB.

  • Private chat assistants

    8B models at full precision or 14B models at 8-bit, behind Open WebUI, for you or your team.

    Ollama guide
  • Image generation

    Stable Diffusion XL and Flux workflows in ComfyUI, on a card nobody else touches.

    GPU images
  • Coding assistants

    A self-hosted code model answering from your own repositories, with no API logging your prompts.

    vLLM guide
  • Rendering and mining

    Blender Cycles renders, and mining, which is allowed on GPU servers.

    Acceptable use

07Offshore, in writing

The rules, before you pay.

What we promise is written into our policies, not just our marketing.

  • US DMCA noticesNot actioned

    Answered with our policy, never enforced. Only a local court order, or a valid EU notice in EU locations, can require action.

    DMCA policy
  • IdentityEmail only

    No name, address, phone or ID document, ever. A private or disposable address is fine.

    Privacy policy
  • Payment5 cryptocurrencies

    Bitcoin, Ethereum, Monero, USDT and Solana, paid on-chain from any wallet to your balance. No card processor, no chargebacks.

    Crypto payments
  • TransparencySigned canary

    A PGP-signed warrant canary every quarter and a public count of every request we receive.

    Warrant canary

08FAQ

RTX 4090 servers, answered.

Another question? The full FAQ answers what people ask before an order: payments, privacy, complaints and support.

Read the full FAQ
Is the RTX 4090 dedicated to me?

Yes. The whole card is yours: no time-slicing, no sharing with other customers, and no spot capacity that can disappear in the middle of a job.

Can I rent it by the hour?

No. Billing is monthly, quarterly (−5%) or yearly (−12%), paid from your balance. At about $0.116 an hour around the clock, a month costs less than hourly rentals once you use the card for a few hours a day.

Which models run well on 24 GB?

8B models at 16-bit and 14B models at 8-bit. A 32B model at 4-bit needs about 24 GB with an 8K context, which leaves no room on this card: the RTX 5090 is the safer pick. For 70B models, choose 2 × RTX 5090 or an A100 80 GB.

How fast is delivery?

Between 1 and 24 hours after you order, depending on the jurisdiction. The access details appear on the server’s page in your client area.

Are the driver and CUDA installed?

Yes, on the Ubuntu 24.04 image. The GPU images guide covers PyTorch, Ollama, vLLM and ComfyUI.

Can I get a refund?

GPU servers are not refundable once delivered, because the hardware is set aside for you. Check the specifications above before you order.

An RTX 4090 of your own, offshore.

  1. 1Pick a jurisdiction
  2. 2Pay in crypto, no ID asked
  3. 3SSH in within 1 to 24 hours
Configure your RTX 4090 All GPU servers

$84.99/mo · dedicated card · CUDA preinstalled

Welcome back

Sign in to manage your servers and your balance.

No KYCHuman check by Cloudflare TurnstileNo tracking