Most deployedCloud GPU
GPU

L4

L4 is available through OneInfer GPU marketplace with live provider pricing, regions, VRAM, and deployment options from the API.

VRAM
24GB
Best price
$0.32/gpu/hr
Providers
7 available
Regions
6 regions
VRAM
24GB
CUDA cores
N/A
TDP
N/A
Bandwidth
N/A
Process node
N/A
Launch date
N/A

Available from 7 providers

Sorted by price · click price to re-sort
ProviderRegionsConfigPrice / hr Availability
vastaivastai
EU-CZ-01, EU-IS-01, EU-CH-01, US-UT-011x 24GB$0.32Limited
runpodrunpod
EU-IS-01, US-MO-021x 24GB$0.49Limited
e2e_networkse2e_networks
IN-DL-011x 24GB$0.57Limited
vastaivastai
EU-CZ-01, EU-IS-012x 24GB$0.67In stock
e2e_networkse2e_networks
IN-DL-012x 24GB$1.14In stock
vastaivastai
EU-IS-014x 24GB$1.34In stock
e2e_networkse2e_networks
IN-DL-014x 24GB$2.29In stock

Full specifications

Data from GPU specs and provider APIs
AI TOPSN/A
FP8 TFLOPSN/A
FP16 TFLOPSN/A
BF16 TFLOPSN/A
FP32 TFLOPSN/A
INT8 TOPSN/A
Sparse multiplierN/A

Benchmarks

L4 vs comparable GPUs on the platform

Inference throughput

7B model · tokens/sec

B200 SXM15,000
H200 SXM8,500
H100620
A100500
GeForce RTX 5090180

Training performance

Overall training multiplier

B200 SXM30x
H200 SXM18x
H10015x
A1003.4x
L40S2.8x

GeekBench OpenCL

Compute benchmark score

GeForce RTX 5090367,740
H200 SXM343,598
L40S337,706
H100336,474
B200 SXM333,768

Best for

LLM fine-tuningHigh-throughput inferenceMulti-modal model trainingLarge-batch distributed trainingRAG pipelines at scale

Similar GPUs

Ready to deploy L4?

Spin up an instance in minutes from your choice of 7 providers.