← Back to all offerings

L4

24 GB VRAM121 FP16 TFLOPS72 W TDP5 providers188 offerings46 regions

The L4 is a cloud GPU with 24 GB of memory and 72 W TDP. It is currently available from 5 providers starting at $0.08/hr, with a market median of $1.62/hr across 188 offerings.

$0.08/hr
Cheapest
on GPU Finder
$1.62/hr
Median
across 188 offerings
$7.60/hr
Most expensive
-25.6%
90-day trend

Best deal

Hardware specifications

Architecture
Ada Lovelace
VRAM
24 GB
Memory bandwidth
300 GB/s
FP16
121 TFLOPS
FP32
30 TFLOPS
TDP
72 W

Same across all providers.

Market comparison

$1.62/hr
Our median / hr
$1.00/hr
gpus.io median / hr
188
Our providers
7
gpus.io providers

View source on gpus.io

All-time price history

Offerings

All ×N×1×2×4×8In stock only · Partners only
ProviderCloudCountvCPURAMRegionPer GPU hrSPOTTotal/hr
HardwareHQ
0.0% 30dReferral link
×2$0.44/hr$0.88/hrLaunch ↗
HardwareHQ
0.0% 30dReferral link
×296 GB$0.95/hr$1.90/hrLaunch ↗
Naver Cloud GPU
0.0% 30dReferral link
×296 GBkr-seoul$1.08/hr$2.15/hrLaunch ↗
Naver Cloud GPU
0.0% 30dReferral link
×296 GBkr-seoul$1.16/hr$2.31/hrLaunch ↗

LLMs that fit

ModelParams16-bit8-bit4-bit
Laguna-S 2.1
Poolside
118B 8B activeNeeds >8 GPUsBarely fits on 8×Barely fits on 5×
gpt-oss-120b
OpenAI
117B 5.1B activeNot releasedNot releasedBarely fits on 4×
Llama 3.1 70B
Meta
70BBarely fits on 7×Barely fits on 4×3× · 18 GB free
Qwen3.6 35B-A3B
Alibaba
35B 3B active5× · 14% spare3× · 15.6 GB free2× · 14.4 GB free
Qwen3 32B
Alibaba
32.8BBarely fits on 4×3× · 20.4 GB free2× · 18 GB free
Qwen2.5 14B
Alibaba
14B2× · 12.6 GB free1× · 4.7 GB free1× · 12.6 GB free
Llama 3.1 8B
Meta
8BBarely fits on 1×1× · 12.6 GB free1× · 17.3 GB free
Mistral 7B
Mistral AI
7B1× · 5.8 GB free1× · 14 GB free1× · 18.1 GB free
Qwen3 4B
Alibaba
4B1× · 12.4 GB free1× · 17.6 GB free1× · 20.3 GB free

Estimates: model weights plus 20% for KV cache and runtime overhead, at short context lengths. Long contexts and large batches need more. Rows marked “tight” hold the weights but not that full margin. A ×N figure is the total VRAM across N of these GPUs and makes no claim about interconnect throughput.

Similar GPUs

Frequently asked questions

How much does the L4 cost per hour?

The cheapest L4 offering starts at $0.08/hr, and the market median is $1.62/hr.

Which cloud providers offer the L4?

We track L4 offerings from HardwareHQ, Naver Cloud GPU.

How much VRAM does the L4 have?

The L4 has 24 GB of VRAM.

What LLMs can run on the L4?

With 24 GB of VRAM, the L4 can run 9 of our tracked models in 4-bit quantization, including Laguna-S 2.1, gpt-oss-120b, Qwen3.6 35B-A3B. 7 models fit in 16-bit precision.

What's the biggest LLM I can run on multiple L4s?

With 8× L4 (192 GB total), the largest tracked model that fits in 16-bit is Qwen3.6 35B-A3B (35B params), which needs 5× GPUs.

Where is the L4 cheapest?

The cheapest L4 offering in our catalog is HardwareHQ at $0.44/hr per GPU-hour (×2).