All 6 data centres operational

Paid in crypto · No identity check · Root access in under 5 minutes

Blackwell · PCIe card

Rent a dedicated RTX 5090 GPU server.

32 GB of GDDR7 at 1,792 GB/s, in a machine that is yours alone. Monthly, paid in crypto, opened without an identity check and delivered in under 5 minutes.

  • $335/moone card, everything included
  • 32 GBGDDR7
  • 1,792 GB/smemory bandwidth
  • 1×, 2×, 4×, 8×node sizes available

Price, by node size and term

One number per line, and it is the whole number: no setup fee, no reinstall fee, no bandwidth charge, no support plan. Longer terms are paid up front.

Monthly price of an RTX 5090 server, by node size and term length
Node CPU and memory Monthly 3 months 12 months
1 × RTX 5090 32 GB total · PCIe 5.0 AMD Ryzen 9 7950X16 c / 32 t · 128 GB · 2 × 2 TB NVMe $335/mo cancel any time $318/mo save $51/yr $295/mo save $480/yr Configure
2 × RTX 5090 64 GB total · PCIe 5.0 ×16 AMD Threadripper 7960X24 c / 48 t · 256 GB · 2 × 3.84 TB NVMe $692/mo cancel any time $657/mo save $105/yr $609/mo save $996/yr Configure
4 × RTX 5090 128 GB total · PCIe 5.0 ×16 AMD Threadripper 7970X32 c / 64 t · 512 GB · 4 × 3.84 TB NVMe $1,345/mo cancel any time $1,278/mo save $201/yr $1,184/mo save $1,932/yr Configure
8 × RTX 5090 256 GB total · PCIe 5.0 ×16 2 × Intel Xeon Gold 6438Y+64 c · 1 TB · 8 × 3.84 TB NVMe $2,584/mo cancel any time $2,455/mo save $387/yr $2,274/mo save $3,720/yr Configure

Every configuration above is available in all 6 data centres, racked and burned in, with an image ready to write — which is why delivery is under 5 minutes rather than a working day.

Full specification

GPUNVIDIA RTX 5090Blackwell architecture, PCIe add-in card.
Memory32 GB GDDR71,792 GB/s of bandwidth — the number that predicts generation speed, far better than any teraflop figure.
InterconnectPCIe 5.0 ×16Peer-to-peer over PCIe. Fine for one model per card and for pipeline parallelism; slower than NVLink for tensor parallelism. Which one your job needs.
Node sizes1 × card, 2 × card, 4 × card, 8 × cardPCIe cards go up to eight in a standard chassis, and ten only on low-power cards.
Storage2 × 2 TB NVMeHanded over raw beyond the system volume — we impose no RAID level. More NVMe and archive HDD are monthly options.
Network1 Gbit/s, unmeteredBoth directions, no included volume and no overage at any traffic level. One IPv4, a routed IPv6 /64, DDoS filtering on from delivery.

What runs on an RTX 5090

Every model below fits on a single card at 8k context — no sharding, no interconnect to think about. The precision shown is the best one that fits; the arithmetic is in the VRAM sizing guide, so you can check it.

Language models that fit on a single RTX 5090 at 8k context
ModelBest precision that fitsMemory neededEstimated tokens/s
Llama 3.1 8B8B parameters BF16 / FP16Reference quality 20 GBof 32 GB ~50
Qwen 3 32B32B parameters 4-bit (AWQ, GPTQ)Smallest weights, cache stays FP16 21 GBof 32 GB ~50
Mistral Small 24B24B parameters FP8Near-reference, Hopper and Blackwell 28 GBof 32 GB ~34
Gemma 3 27B27B parameters 4-bit (AWQ, GPTQ)Smallest weights, cache stays FP16 20 GBof 32 GB ~60
Phi-4 14B14B parameters FP8Near-reference, Hopper and Blackwell 17 GBof 32 GB ~58

On the 8-card node the same models are held across 256 GB in total, which changes what is possible entirely — up to Llama 3.1 405B. Sharding costs synchronisation at every layer, so a model that fits on one card should stay on one card.

Throughput to expect

Generation is bounded by memory bandwidth: each token requires re-reading the active weights. At 1,792 GB/s, that sets a ceiling no amount of tuning gets past.

Our figures are a floor, not a promise. The estimates on this page are calibrated against published measurements and are deliberately conservative — they model single-stream generation. With continuous batching under vLLM the aggregate across concurrent requests is several times higher. If a number here looks low against a benchmark you have seen, that is usually the difference.

Compared with nearby cards

The three cards closest to this one in price. Memory decides what loads; bandwidth decides how fast it answers.

The RTX 5090 compared with the three closest cards by price
CardMemoryBandwidthInterconnectFrom
RTX 5090 This pageBlackwell 32 GB 1,792 GB/s PCIe $335/mo
RTX A6000Ampere 48 GB 768 GB/s PCIe $286/mo
A100 PCIeAmpere 40 GB 1,555 GB/s PCIe $392/mo
RTX 4090Ada Lovelace 24 GB 1,008 GB/s PCIe $193/mo

See all 12 cards side by side, or let the configurator pick from your model and context length instead of from a price.

Compared with other cards

The arbitrations people actually make against an RTX 5090, each worked out on price, memory, bandwidth and the models that fit.

What is included

  • No identity check

    No document, no selfie, no phone number, no company registration.

    Nothing verified · at any spend

  • The whole card

    The GPU is passed through to one machine, and that machine is yours.

    No MIG · no vGPU · no hypervisor

  • Root and IPMI

    Full root from the first minute, plus remote power and virtual media.

    Out-of-band on its own VLAN

  • Unmetered bandwidth

    No included volume, no overage tier, nothing to watch on a graph.

    1 to 25 Gbit/s · both directions

  • No setup or reinstall fee

    Reimage as often as you like, to any operating system we offer.

    $0 · reinstalls unlimited

  • Hardware swapped, fast

    From spares held in the same suite, at any hour, disks left in place.

    Four-hour target · 24/7

Where you can have it

Every RTX 5090 configuration is available in all 6 data centres. There is no site where a card costs more, and none where it is unavailable.

AMS Available

Amsterdam

Netherlands

Round trip7 ms to Frankfurt, 12 ms to London
FacilityTier III · 100% wind
STO Available

Stockholm

Sweden

Round trip22 ms to Frankfurt, 31 ms to London
FacilityTier III · 100% hydro
ZRH Available

Zürich

Switzerland

Round trip9 ms to Milan, 15 ms to Frankfurt
FacilityTier IV · 96% carbon-free
REY Available

Reykjavík

Iceland

Round trip18 ms to London, 40 ms to New York
FacilityTier III · 100% geothermal + hydro
YUL Available

Montréal

Canada

Round trip12 ms to New York, 19 ms to Toronto
FacilityTier III · 99% hydro
SIN Available

Singapore

Singapore

Round trip38 ms to Tokyo, 45 ms to Sydney
FacilityTier III · Grid + REC offset

RTX 5090 server questions

How much does an RTX 5090 server cost per month?
A single-card RTX 5090 server is $335 a month, with everything included: 128 GB of system RAM, 2 × 2 TB NVMe, unmetered 1 Gbit/s, one IPv4 address and a routed IPv6 /64. There is no setup fee and no bandwidth charge. Multi-GPU nodes run from $692 for 2 cards up to $2,584 for 8.
How much VRAM does the RTX 5090 have, and what fits in it?
32 GB of GDDR7 per card, at 1,792 GB/s. The largest model it holds on one card at 8k context is Qwen 3 32B (4-bit (AWQ, GPTQ)). A 8-card node has 256 GB in total, enough for Llama 3.1 405B.
Is the RTX 5090 dedicated, or shared with other customers?
Dedicated. The card is passed through to a single physical machine that only you have an account on — no MIG partitioning, no vGPU layer and no other tenant on the same silicon. On a multi-GPU node you get the whole node, including the PCIe 5.0 ×16 fabric between the cards.
Do I need to verify my identity to rent an RTX 5090?
No. There is no document upload, no selfie, no phone number and no company registration, at any level of spend — an eight-card node opens on exactly the same terms as the cheapest single card. Opening an account takes an email address, a password and a billing address, and none of it is verified.
How do I pay, and how quickly is the server delivered?
In cryptocurrency only: Bitcoin, Monero, Ethereum (ERC-20), Tether (ERC-20) and others. Each invoice gets its own address, the amount is fixed at the rate shown when it is issued, and the machine is provisioned on the first confirmation — root access in under 5 minutes, at any hour.
Can I rent an RTX 5090 by the hour instead?
Not here. Every price on this site is a monthly price, with no meter and no per-second billing. Against a typical $2.99/hour cloud, a monthly term is cheaper past roughly 113 hours of the machine simply existing — about 5 days out of thirty.

An RTX 5090 of your own, from $335 a month.

Racked, burned in and waiting in all 6 data centres. Pay in crypto and it is yours in under 5 minutes.

Sign in

Console, invoices and out-of-band access.

No account yet?

There is no separate sign-up. Your account is created while you place your first order — you choose the email and the password on the payment step, and the console is open by the time the machine is.

Configure a server

Language