L40S, dedicated.
A single-GPU L40S machine, 48 GB of GDDR6 on NVIDIA's Ada Lovelace architecture, on an x86 host. Listed at $0.79 per hour, billed per minute, the same terms as every other GPUwerk tier.
No capacity available right now. Join the waitlist: when a machine is free, we hold it for the first person in line with at least one hour of credit and email them. This page describes a tier we are testing demand for, not hardware currently installed in the fleet.
Specifications
Per node. Every GPUwerk instance, on any tier, is one whole machine to yourself.
- GPU
- L40S, Ada Lovelace architecture
- Memory
- 48 GB GDDR6
- Host
- x86 host configuration listed in the console
- Tenancy
- One customer per machine, never shared
We have not measured tok/s or training throughput on this tier yet, since no unit is in the fleet. We publish figures only after we run them ourselves; see the DGX Spark fleet benchmarks for how we report them once we do.
What would you run on it?
48 GB positions the L40S between the 5090's 32 GB and a Spark's 128 GB, with data-center-class ECC memory.
Production inference
A data-center card built for sustained serving load, where ECC memory and driver support matter more than a consumer part.
Larger single-model serving
A 30B to 70B-class model at Q4 fits with headroom the 5090 does not have.
Rendering and graphics workloads
The L40S carries RT cores and a media engine the H100 and B200 lack, useful if your pipeline mixes graphics with inference.
Billing
- Rate
- $0.79 per hour, listed
- Billing
- Per minute, against a prepaid credit balance
- Fees
- No egress fees
- Commitment
- None
Prices are in USD, excluding applicable tax. Full billing mechanics, including what a stop versus a terminate does, are on the pricing page.
FAQ
Is the L40S tier available now?
No capacity is available right now. We list the tier and its rate so you can join the waitlist; if a machine becomes free, we hold it for the first person in line who has at least one hour of credit, and email them. Nothing on this page claims an L40S is installed or booked today.
What's the host configuration around the GPU?
The host configuration (CPU, system RAM, storage) is listed in the console at deploy time, once the tier has capacity. We are not publishing numbers for a machine that is not yet in the fleet.
How does billing work on this tier?
The same as every GPUwerk tier: per minute, against a prepaid credit balance, in USD, excluding applicable tax, with no egress fees and no commitment.
Weighing capacity against raw VRAM?
See how a Spark's 128 GB compares to a discrete card, or check current rates on the pricing page.
Join the waitlist