Vera Rubin
UK
{"@context":"https://schema.org","@graph":[{"@type":"Service","@id":"https://gpuaas.com/gpu/vera-rubin-llm-training#service","name":"Vera Rubin for LLM Training","provider":{"@type":"Organization","name":"GPUaaS.com","url":"https://gpuaas.com"},"serviceType":"GPU cloud infrastructure","description":"Vera Rubin NVL72 for LLM training: 72 Rubin GPUs in one NVLink 6 domain, specs, early-access price and how it compares with GB300. Register interest."},{"@type":"WebPage","@id":"https://gpuaas.com/gpu/vera-rubin-llm-training#webpage","url":"https://gpuaas.com/gpu/vera-rubin-llm-training","name":"Vera Rubin for LLM Training","isPartOf":{"@type":"WebSite","name":"GPUaaS.com","url":"https://gpuaas.com"}},{"@type":"BreadcrumbList","@id":"https://gpuaas.com/gpu/vera-rubin-llm-training#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https://gpuaas.com"},{"@type":"ListItem","position":2,"name":"GPU Cloud","item":"https://gpuaas.com/cluster"},{"@type":"ListItem","position":3,"name":"Vera Rubin for LLM Training","item":"https://gpuaas.com/gpu/vera-rubin-llm-training"}]},{"@type":"FAQPage","@id":"https://gpuaas.com/gpu/vera-rubin-llm-training#faq","mainEntity":[{"@type":"Question","name":"Is Vera Rubin worth waiting for over GB300 for training?","acceptedAnswer":{"@type":"Answer","text":"Only for frontier-scale runs that can wait for reservation-only access and are large enough to need one NVLink domain across 72 GPUs. GB300 is the same rack-scale design available now, so for a training run starting this year, GB300 is the more realistic choice."}},{"@type":"Question","name":"What are the Vera Rubin NVL72 specs for training?","acceptedAnswer":{"@type":"Answer","text":"Vera Rubin NVL72 pairs 72 Rubin GPUs with 36 Vera CPUs (88 Arm Olympus cores each), 288GB of HBM4 per GPU at 22TB/s, connected via NVLink 6, with an estimated rack draw of 190-230kW. Some providers brand the Rubin GPU as \"H300\"; it is the same silicon regardless of name."}},{"@type":"Question","name":"How does Vera Rubin compare with GB300 for training?","acceptedAnswer":{"@type":"Answer","text":"Both are 72-GPU rack-scale NVL72 designs. Vera Rubin brings the next GPU generation, HBM4 memory and NVLink 6, at a higher estimated rack draw. Whether that shortens your run enough to justify waiting depends on your model and parallelism strategy, and published figures are not a substitute for testing your own pipeline."}},{"@type":"Question","name":"What is Vera Rubin availability for training today?","acceptedAnswer":{"@type":"Answer","text":"Broad on-demand access is not available yet. Vera Rubin is on early-access reservation terms through a small set of cloud partners, and confirmed regional availability varies by market. Register interest with your timeline and we will confirm what is securable."}},{"@type":"Question","name":"Will my training frameworks work on Vera Rubin?","acceptedAnswer":{"@type":"Answer","text":"Existing parallelism strategies and frameworks such as PyTorch FSDP, DeepSpeed and Megatron-LM are the starting point, but framework and library support for a new generation matures over time. Confirm support for your exact stack before planning a run around Vera Rubin."}},{"@type":"Question","name":"What is the Vera Rubin price per hour for training?","acceptedAnswer":{"@type":"Answer","text":"Tracked early-access Vera Rubin pricing starts around $11.00/hr on a reservation-only basis as of September 2026, with one July 2026 industry analysis citing roughly $8.50/hr per chip in 3-year rental cost comparisons. The market is very thin, so Vera Rubin pricing through GPUaaS.com is quoted per enquiry and varies by commitment term and configuration."}}]}]}
GPUAAS.COM · WHOLESALE GPU NETWORK / A HOSTED·AI SERVICE ◆ CAPACITY AVAILABLE · 20+ PARTNERSQUOTES < 24HREV 2026.09
+
+
◆
Vera Rubin NVL72 for LLM training
◆ AVAILABLE

Vera Rubin
for LLM training
, at
wholesale price.

Register interest in Vera Rubin NVL72 early access, for frontier-scale training across one rack-wide NVLink 6 domain, at
~30% less than hyperscale. Quotes in under 24 hours.

HGX GPU node
GPU generations
8
Architectures
Hopper + Blackwell + Vera Rubin
Vetted partners
20+
Quote turnaround
24 hrs
Commitment
Short / long
QUOTES IN UNDER 24 HOURS VETTED PARTNERS WORLDWIDE SHORT OR LONG TERM COMMITMENT DIRECT OPERATOR CONTRACTS CAPACITY AVAILABLE NOW PLACEMENT YOU SPECIFY
◆ THE SHORT ANSWER

Vera Rubin NVL72 (VR200) puts 72 Rubin GPUs and 36 Vera CPUs in a single NVLink 6 domain, with 288GB of HBM4 per GPU. For training, the draw is the same one rack-scale designs offer, keeping tensor and expert-parallel traffic inside the rack instead of crossing slower node-to-node networking, with a new generation of memory bandwidth behind it. Access is early and reservation-only today, with tracked Vera Rubin pricing from roughly $11.00/hr in a very thin market, so it suits frontier-scale runs planned well ahead. Register interest and we will confirm what is actually securable. GB300 for LLM training, B300 and B200 are the options available for training capacity now.

+
01
PRICING

What Vera Rubin training is expected to cost

Tracked early-access Vera Rubin pricing starts around $11.00/hr on a reservation-only basis, in a very thin market. Vera Rubin pricing is quoted per enquiry; full Vera Rubin NVL72 specs are available on request.

Market reference as of September 2026, quoted in USD. This is an early, thinly tracked market with few providers quoting publicly, so figures are directional. Training cost depends on cluster size, parallelism and run length, so total cost per run is the number to calculate.
Wholesale rates through GPUaaS.com are quoted per enquiry and vary by commitment term, configuration and placement.

$0$2.50$5$7.50$10$12.50$15/GPU-HR
Early-access rate, reservation-only
Tracked provider, reservation-only
$11.00
Industry reference (3-yr rental TCO)
Per-chip figure, July 2026 analysis
$8.50
Approx. premium band, high end
Directional, thin provider set
~$16.50
Hyperscaler on-demand (projected)
Likely floor once hyperscalers quote broadly
~$18.00
◆ GPUaaS.com wholesale
Vetted partners · direct operator contract
quoted per enquiry
◆ VERY THIN, EARLY MARKET: FEW PROVIDERS QUOTE YET
+
02
◆
Where Vera Rubin is expected to matter for training

What Vera Rubin should handle well for training, and what to watch for.

Vera Rubin NVL72 is the next rack-scale generation after Blackwell Ultra: 72 Rubin GPUs, 36 Vera CPUs, 288GB of HBM4 per GPU and one NVLink 6 domain. For training, the value is the same one rack-scale designs offer, keeping parallelism strategies that would normally cross node boundaries on the fastest interconnect, now with HBM4 behind it. It is also early. Access is reservation-only, supply is thin, an estimated 190-230kW per rack sits above GB300's roughly 135-140kW, and framework support for a new generation matures over time. For a run that has to start this year, GB300 and B300 are the realistic options, and Vera Rubin is a plan for frontier-scale training that can be scheduled well ahead.

/01

A rack-wide NVLink 6 domain

72 Rubin GPUs share one NVLink 6 domain, keeping tensor and expert-parallel traffic off slower node-to-node networking.
NVLink 6 · 72 GPUs · one domain
/02

HBM4 memory per GPU

288GB of HBM4 per GPU, at a new generation of bandwidth, gives larger models and batches room inside the domain.
288GB HBM4 · 22TB/s · larger batches
/03

Early access, narrow supply

Reservation-only access today means Vera Rubin fits runs planned well ahead, not training that needs to start now.
reservation-only · plan ahead · early access
/04

Power, cooling and lead times

An estimated 190-230kW per rack means liquid cooling and secured power, and lead times need confirming.
190-230kW · liquid cooling · lead times
+
03
◆ LIVE NETWORK · 12 LOCATIONS

Vera Rubin capacity worldwide, in the location you need.

Vera Rubin training capacity is placement-sensitive and still forming: confirm real in-country supply and lead times before planning a run. See Vera Rubin availability by country below.

Read the full guide to GPU cloud in this location →
8
GPU GENERATIONS
20+
VETTED PARTNERS
12
PLACEMENT OPTIONS
24h
QUOTE TURNAROUND
◆ USA◆ CAN◆ UK◆ DEU◆ FRA◆ NLD◆ UAE◆ SAU◆ IND◆ SGP◆ JPN◆ AUS

Every location links to its own page. Click through for local pricing and specs.

04
◆ COST COMPARISON

See how much you save at scale

Wholesale rates against cloud list price for a 64-GPU cluster.

CLUSTER SIZE
8 GPU Servers
64 × GPUS · 730 HRS/MO
ASSUMPTIONS · BLENDED $6.00/GPU-HR · INDICATIVE ONLY
SOURCEEST. MONTHLYVS GPUAAS
Retail cloud
On-demand list price · reserved discounts require lock-in
~$280k
+$84k
Direct datacentre negotiation
Long-term commitment · slow procurement cycle
~$230k
+$34k
◆ BEST VALUE
GPUaaS.com wholesale
Vetted partners · direct operator contract · quotes in 24 hours
~$196k
SAVE ~$84k/MO
Need single-GPU compute? packet.ai has you covered.
+
05
◆ HOW IT WORKS

A matchmaker, not a marketplace.

We connect you to our vetted partners. You contract directly with the operator running your nodes.

STEP 01/4
01

Tell us the requirement

GPU model, count, placement and timeline. Add workload detail if you have it.

STEP 02/4
02

We match capacity

We find vetted partners with capacity that fits, in the jurisdiction you need.

STEP 03/4
03

Quotes in 24 hours

Real quotes from partners who hold the capacity, not listings that may not exist.

STEP 04/4
04

Contract and provision

You contract directly with the operator. We smooth the provisioning process.

Get a quote
Request wholesale rates
in under 24 hours.

Tell us the essentials. We'll line up real quotes from our vetted wholesale partners, and you contract directly with the operator.

◆Quotes in under 24 hours
◆Direct contact with operators
◆Vetted partners, matched to your requirement
◆20+ vetted providers · 12 locations
1
ESSENTIALS
2
OPTIONAL
Contact
Full Name *
Business Email *
Organization *
Preferred Location *
Your Region *
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.
GPU Requirements
GPU Model *
PRE-SELECTED
Number of GPUs *
Individual GPU count. 1 node = 8 GPUs.
Get the Best Deal→
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.
+
06
◆ FAQ

Frequently Asked Questions

Q1
Is Vera Rubin worth waiting for over GB300 for training?

Only for frontier-scale runs that can wait for reservation-only access and are large enough to need one NVLink domain across 72 GPUs. GB300 is the same rack-scale design available now, so for a training run starting this year, GB300 is the more realistic choice.

Q2
What are the Vera Rubin NVL72 specs for training?

Vera Rubin NVL72 pairs 72 Rubin GPUs with 36 Vera CPUs (88 Arm Olympus cores each), 288GB of HBM4 per GPU at 22TB/s, connected via NVLink 6, with an estimated rack draw of 190-230kW. Some providers brand the Rubin GPU as "H300"; it is the same silicon regardless of name.

Q3
How does Vera Rubin compare with GB300 for training?

Both are 72-GPU rack-scale NVL72 designs. Vera Rubin brings the next GPU generation, HBM4 memory and NVLink 6, at a higher estimated rack draw. Whether that shortens your run enough to justify waiting depends on your model and parallelism strategy, and published figures are not a substitute for testing your own pipeline.

Q4
What is Vera Rubin availability for training today?

Broad on-demand access is not available yet. Vera Rubin is on early-access reservation terms through a small set of cloud partners, and confirmed regional availability varies by market. Register interest with your timeline and we will confirm what is securable.

Q5
Will my training frameworks work on Vera Rubin?

Existing parallelism strategies and frameworks such as PyTorch FSDP, DeepSpeed and Megatron-LM are the starting point, but framework and library support for a new generation matures over time. Confirm support for your exact stack before planning a run around Vera Rubin.

Q6
What is the Vera Rubin price per hour for training?

Tracked early-access Vera Rubin pricing starts around $11.00/hr on a reservation-only basis as of September 2026, with one July 2026 industry analysis citing roughly $8.50/hr per chip in 3-year rental cost comparisons. The market is very thin, so Vera Rubin pricing through GPUaaS.com is quoted per enquiry and varies by commitment term and configuration.