L40S, dedicated.

A single-GPU L40S machine, 48 GB of GDDR6 on NVIDIA's Ada Lovelace architecture, on an x86 host. Listed at $0.79 per hour, billed per minute, the same terms as every other GPUwerk tier.

No capacity available right now. Join the waitlist: when a machine is free, we hold it for the first person in line with at least one hour of credit and email them. This page describes a tier we are testing demand for, not hardware currently installed in the fleet.

Specifications

Per node. Every GPUwerk instance, on any tier, is one whole machine to yourself.

NVIDIA L40S
GPU
L40S, Ada Lovelace architecture
Memory
48 GB GDDR6
Host
x86 host configuration listed in the console
Tenancy
One customer per machine, never shared

We have not measured tok/s or training throughput on this tier yet, since no unit is in the fleet. We publish figures only after we run them ourselves; see the DGX Spark fleet benchmarks for how we report them once we do.

What would you run on it?

48 GB positions the L40S between the 5090's 32 GB and a Spark's 128 GB, with data-center-class ECC memory.

Production inference

A data-center card built for sustained serving load, where ECC memory and driver support matter more than a consumer part.

Larger single-model serving

A 30B to 70B-class model at Q4 fits with headroom the 5090 does not have.

Rendering and graphics workloads

The L40S carries RT cores and a media engine the H100 and B200 lack, useful if your pipeline mixes graphics with inference.

Billing

Rate
$0.79 per hour, listed
Billing
Per minute, against a prepaid credit balance
Fees
No egress fees
Commitment
None

Prices are in USD, excluding applicable tax. Full billing mechanics, including what a stop versus a terminate does, are on the pricing page.

FAQ

Is the L40S tier available now?

No capacity is available right now. We list the tier and its rate so you can join the waitlist; if a machine becomes free, we hold it for the first person in line who has at least one hour of credit, and email them. Nothing on this page claims an L40S is installed or booked today.

What's the host configuration around the GPU?

The host configuration (CPU, system RAM, storage) is listed in the console at deploy time, once the tier has capacity. We are not publishing numbers for a machine that is not yet in the fleet.

How does billing work on this tier?

The same as every GPUwerk tier: per minute, against a prepaid credit balance, in USD, excluding applicable tax, with no egress fees and no commitment.

Weighing capacity against raw VRAM?

See how a Spark's 128 GB compares to a discrete card, or check current rates on the pricing page.

Join the waitlist