Glamdring

Model infrastructure

Lambda Labs pricing: Lambda price list, September 2026

For H100 and B200, Lambda's 8-GPU instance is its cheapest published rate ($3.99 and $6.69 per GPU-hour, before tax, checked 28 September 2026). 1-Click Clusters cost 32.6% to 54.4% more per GPU-hour and are priced by GPU count, not term, across 2 weeks to 1 year.

Glamdring Research13 min6 sources checked

Engraving of a tower GPU workstation standing beside a rack-mounted GPU cluster node

Lambda’s cheapest GPU instance is a single NVIDIA Quadro RTX 6000 at $0.69 per GPU-hour, on Lambda’s pricing page checked 28 September 2026.1 The dearest is a single NVIDIA B200 SXM6 at $6.99, and an 8-GPU H100 SXM instance costs $3.99 per GPU-hour.1 1-Click Clusters cost $5.54 to $6.16 per H100-hour and $8.87 to $9.86 per B200-hour on terms of 2 weeks to 1 year.1 Every price is before sales tax, VAT or GST.1

TL;DR

  • The 8-GPU instance is the cheapest published way to buy an H100 or B200 hour at Lambda: $3.99 and $6.69 per GPU-hour. Each smaller size adds $0.10, up to $4.29 and $6.99 for one GPU.1
  • The A100 40 GB costs $1.99 per GPU-hour at every instance size, and the A6000 costs $1.09 at every size it’s sold in.1
  • Cluster prices fall with GPU count, not with term. Across the 2-week-to-1-year band the rate is fixed; 256 GPUs cost 10% less per GPU-hour than 16.1
  • A 16-GPU H100 cluster costs 54.4% more per GPU-hour than an 8x H100 instance. For the B200 the gap is 47.4%.1
  • Terms beyond one year and Superclusters carry no public price. Lambda routes its lowest prices, for reserved capacity, through its own team.1
  • At least one third-party tracker lists an A100 price that Lambda’s page doesn’t show. Where they differ, use Lambda’s page.2,1

How this ranking was built

This page records Lambda’s published instance and cluster prices per GPU-hour, checked 28 September 2026. It sits in our model infrastructure category. The source is one page: Lambda’s AI cloud pricing page, captured at 15:05 Australian Eastern Standard Time, which shows no last-updated date.1

The ranked field is Lambda’s PRICE/GPU/HR column in its Instances pricing tables.1 The page shows four instance tables under the tabs 8x, 4x, 2x and 1x. The capture doesn’t label each table, so we matched them by order and by their resources, which halve from one table to the next: a B200 instance has 208 vCPUs in the first table and 26 in the fourth.1 ComputePrices, a price tracker, also labels the $6.69 B200 and $3.99 H100 rows as 8-GPU configurations.2

The four tables hold 22 rows. All 22 carry a price and all are ranked, cheapest first; tied prices share a rank. The 1-Click Clusters tables hold 8 rows, 6 of them priced, and sit in their own table below because they are priced by GPU count and term.1 Hourly totals multiply the per-GPU rate by the GPU count. Percentages are our arithmetic, rounded to one decimal place. Our research methodology sets out how prices are captured and checked.

What does each Lambda GPU instance cost per GPU-hour?

Lambda’s instance prices run from $0.69 per GPU-hour for a single Quadro RTX 6000 to $6.99 for a single B200 SXM6, on the pricing page checked 28 September 2026.1 The last column is our product of the listed rate and the GPU count.

RankGPUVRAM per GPUGPUsPrice per GPU-hour ($, before tax)Instance per hour ($, computed)
1NVIDIA Quadro RTX 600024 GB1$0.69$0.69
2NVIDIA Tesla V10016 GB8$0.79$6.32
=3NVIDIA A600048 GB4$1.09$4.36
=3NVIDIA A600048 GB2$1.09$2.18
=3NVIDIA A600048 GB1$1.09$1.09
6NVIDIA A1024 GB1$1.29$1.29
=7NVIDIA A100 SXM40 GB8$1.99$15.92
=7NVIDIA A100 PCIe40 GB4$1.99$7.96
=7NVIDIA A100 PCIe40 GB2$1.99$3.98
=7NVIDIA A100 SXM40 GB1$1.99$1.99
=7NVIDIA A100 PCIe40 GB1$1.99$1.99
12NVIDIA GH20096 GB1$2.29$2.29
13NVIDIA A100 SXM80 GB8$2.79$22.32
14NVIDIA H100 PCIe80 GB1$3.29$3.29
15NVIDIA H100 SXM80 GB8$3.99$31.92
16NVIDIA H100 SXM80 GB4$4.09$16.36
17NVIDIA H100 SXM80 GB2$4.19$8.38
18NVIDIA H100 SXM80 GB1$4.29$4.29
19NVIDIA B200 SXM6180 GB8$6.69$53.52
20NVIDIA B200 SXM6180 GB4$6.79$27.16
21NVIDIA B200 SXM6180 GB2$6.89$13.78
22NVIDIA B200 SXM6180 GB1$6.99$6.99

The ranking sorts by the rate, not by the bill. An 8x B200 instance is the dearest thing to run by the hour at $53.52, yet its per-GPU rate is the cheapest B200 on the page.1

Several GPUs appear at one size only. The V100 and the 80 GB A100 come only as 8-GPU instances, and the GH200, A10, H100 PCIe and Quadro RTX 6000 only as single GPUs.1 That changes the entry cost more than the rate does. The cheapest way onto an 80 GB A100 is $22.32 an hour for eight of them, while an H100 PCIe with the same 80 GB of memory costs $3.29 an hour on its own.1

Our explainer on what controls AI inference cost covers the other inputs to a serving bill, from utilisation to the hardware choice itself.

How does instance size change Lambda’s price?

On Lambda’s two newest instance GPUs, each halving of instance size adds $0.10 per GPU-hour. The H100 SXM runs $3.99, $4.09, $4.19 and $4.29 from 8 GPUs down to 1. The B200 SXM6 runs $6.69, $6.79, $6.89 and $6.99.1

The same $0.30 step from 8 GPUs to 1 weighs differently on each. It is 7.5% of the 8x H100 rate and 4.5% of the 8x B200 rate.1 Older GPUs carry no size premium at all: the A100 40 GB is $1.99 at every size and the A6000 is $1.09 at 4, 2 and 1 GPUs.1

The resources attached to each GPU stay close to constant as the instance shrinks. An 8x B200 instance lists 208 vCPUs, 2,900 GiB of RAM and 22 TiB of SSD; the single B200 lists 26 vCPUs, 360 GiB and 2.75 TiB.1 Per GPU, that’s 26 vCPUs either way and 362.5 GiB against 360 GiB of RAM. So the small-instance premium buys the ability to rent fewer GPUs, not a thinner machine per GPU.

The decision changes at the single-GPU tier. There, the H100 PCIe at $3.29 is $1.00 an hour cheaper than the H100 SXM at $4.29, with the same 80 GB listed and 1 TiB of SSD against 2.75 TiB.1 A single-GPU job that doesn’t need the SXM part saves 23.3% by taking the PCIe card.

What do Lambda’s 1-Click Clusters cost by term?

Lambda’s 1-Click Clusters cost $5.54 to $6.16 per H100 GPU-hour and $8.87 to $9.86 per B200 GPU-hour for any term from 2 weeks to 1 year, on the pricing page checked 28 September 2026.1 The rate falls as the cluster grows; the term inside that band doesn’t move it. Clusters of a year or more show no price.

GPUTermGPUsPrice per GPU-hour ($, before tax)Cluster per hour ($, computed)
NVIDIA HGX B2002 weeks – 1 year16$9.86$157.76
NVIDIA HGX B2002 weeks – 1 year64$9.36$599.04
NVIDIA HGX B2002 weeks – 1 year256+$8.87$2,270.72 at 256
NVIDIA HGX B2001 year+16+No price; talk to LambdaNot listed
NVIDIA H1002 weeks – 1 year16$6.16$98.56
NVIDIA H1002 weeks – 1 year64$5.85$374.40
NVIDIA H1002 weeks – 1 year256$5.54$1,418.24
NVIDIA H1001 year+16+No price; talk to LambdaNot listed

Lambda describes these as production-ready clusters of 16 to 2,000+ B200 or H100 GPUs.1 Every priced row still routes to “Talk to our team”; nothing on the cluster tables is self-serve checkout.1

The volume steps are almost identical on both GPUs. Moving from 16 to 64 GPUs cuts the rate by 5.0% on the H100 and 5.1% on the B200. Moving from 16 to 256 cuts it by 10.1% on the H100 and 10.0% on the B200.1

In cash terms, the smallest cluster for the shortest term is large. At 336 hours for two weeks of continuous use, 16 H100s cost $33,116.16 and 16 B200s cost $53,007.36 before tax.1 The estimate assumes the whole term bills at the listed rate. The page doesn’t state how a term is invoiced, so confirm it in the current quote.

Above one year, the page stops publishing prices. It says to contact Lambda “for reserved capacity at our lowest prices”, which implies the published cluster rates aren’t the floor.1 Superclusters, the third product named at the top of the page, have no price table in the capture.1

How much more does a Lambda cluster cost than an instance?

A 1-Click Cluster costs 32.6% to 54.4% more per GPU-hour than Lambda’s 8-GPU instance of the same GPU. The smallest H100 cluster, at $6.16, sits $2.17 above the 8x H100 instance at $3.99. The smallest B200 cluster, at $9.86, sits $3.17 above the 8x B200 at $6.69.1

GPU8x instance16-GPU clusterPremium256-GPU clusterPremium
H100$3.99$6.1654.4%$5.5438.8%
B200$6.69$9.8647.4%$8.8732.6%

Take 16 H100s for two weeks. Two 8x instances cost $63.84 an hour, or $21,450.24 over 336 hours. The 16-GPU cluster costs $98.56 an hour, or $33,116.16.1 The difference is $11,665.92.

The page doesn’t itemise what the premium covers. What it does show is the access model. Instances are “self-serve, first-come access”; clusters are production-ready blocks sold for a term through Lambda’s team.1 One tracker describes the clusters as built for distributed workloads.3 The constraint is availability. On 28 September 2026 another tracker marked Lambda’s cheapest listed B200, GH200 and V100 configurations as sold out.4 A buyer who can’t wait for first-come capacity pays the cluster rate for a committed block. A buyer whose job fits on eight GPUs and can wait pays the instance rate.

Is sales tax included in Lambda’s GPU prices?

No. Every price table on Lambda’s page carries the footnote “plus applicable sales tax/VAT/GST”, including all four instance tables and both cluster tables.1 The listed rate is the pre-tax rate.

The footnote names three tax types and no rate, so the tax added depends on where the buyer is billed.1 The page also prints a dollar sign without a currency code.1 Treat both as open items in any budget: confirm the invoicing currency and the tax Lambda will add for your billing entity in the current quote. Every figure on this page, including our hourly and two-week totals, is before tax.

Why do other published Lambda prices disagree?

Other published Lambda prices disagree for three traceable reasons: a tracker’s row that doesn’t match Lambda’s page, old reservation figures still ranking in search, and billing details that Lambda’s page doesn’t state.

ComputePrices lists the A100 SXM 80 GB at $1.99 per hour.2 Lambda’s page lists the A100 SXM 80 GB at $2.79 per GPU-hour, in the 8x table only; $1.99 is its rate for the 40 GB A100.1 The same tracker’s B200, GH200, H100 PCIe, H100 SXM, A10 and A6000 rates match Lambda’s page.2,1

The Next Platform’s February 2024 report on Lambda’s $320 million raise said Lambda and CoreWeave advertised GPU pricing as low as $2.23 per GPU per hour, usually for one-to-three-year reservations.5 That figure is more than two and a half years old and still appears in search results for Lambda’s pricing. Lambda’s current page publishes no price for any term over one year.1

Billing increments differ by source. Fluence, a GPU marketplace that sells against Lambda, describes Lambda as billing per minute.6 Lambda’s captured pricing page states no billing increment.1 On a short job the increment decides the bill, so check it on a small first run.

What the ranking cannot tell you

This ranking shows Lambda’s list prices on 28 September 2026. They are a company statement about what Lambda charges. They don’t show what reserved or enterprise customers negotiate, and Lambda says its lowest prices are for reserved capacity.1

The page shows price, not availability. A GPU on the list may be sold out on the day you need it; one tracker flagged Lambda’s cheapest B200, GH200 and V100 configurations as sold out on the check date.4 The ranking doesn’t measure performance per dollar, so a cheaper GPU-hour isn’t a cheaper job.

Storage and data transfer sit outside the captured page. Trackers list Lambda block storage at $0.20 per GB per month and state that Lambda charges no egress fees.4,3 Confirm both before you model a data-heavy workload. We have no earlier capture of this page, so the tables show no price changes. The ranking will be rebuilt when Lambda changes its pricing page. New pricing research goes out in our free research briefing.

Frequently asked questions

How much does an A100 cost per hour on Lambda?

A 40 GB A100 costs $1.99 per GPU-hour on Lambda at every size, from one GPU to eight, on the pricing page checked 28 September 2026.1 The 80 GB A100 SXM costs $2.79 per GPU-hour and comes only as an 8-GPU instance, which is $22.32 an hour.1 Both are before tax.

What is the cheapest GPU on Lambda?

The NVIDIA Quadro RTX 6000 is the cheapest, at $0.69 per GPU-hour for a single-GPU instance with 24 GB of memory.1 The Tesla V100 is second at $0.79 per GPU-hour, but it’s sold only as an 8-GPU instance, so the hourly bill is $6.32.1 For more memory on one GPU, a single A6000 gives 48 GB for $1.09 an hour.1

Does Lambda publish prices for commitments longer than a year?

No. Both cluster tables list a “1 year+” row for 16 or more GPUs with a dash in the price column and a link to talk to Lambda’s team.1 The page says reserved capacity at Lambda’s lowest prices is available by contacting the company. Superclusters also carry no public price.1

Is Lambda Labs legit?

Lambda, formerly Lambda Labs, was founded in 2012.5 In February 2024 it raised a $320 million Series C round, which The Next Platform reported brought its total raised to $432 million.5 Lambda states SOC 2 Type II and ISO 27001 on its trust site, according to a tracker that links to it.4 Those are company statements and funding records, not an audit of service quality. Check the scope of any attestation before you rely on it.

Sources checked

  1. AI cloud pricingChecked September 28, 2026
  2. Lambda Labs vs VoltageGPUChecked September 28, 2026

Company-owned pages establish what a company says. They do not prove a market conclusion. Each source is dated so readers can judge each claim.