← Back to all offerings

RTX 5070 Ti

16 GB VRAM300 W TDP2 providers6 offerings3 regions

The RTX 5070 Ti is a cloud GPU with 16 GB of memory and 300 W TDP. It is currently available from 2 providers starting at $0.12/hr, with a market median of $0.17/hr across 6 offerings.

$0.12/hr
Cheapest
on Vast.ai
$0.17/hr
Median
across 6 offerings
$0.64/hr
Most expensive
-18.8%
90-day trend

Best deal

Hardware specifications

Architecture
Blackwell
VRAM
16 GB
Memory type
GDDR7
Memory bandwidth
896 GB/s
FP32
41 TFLOPS
TDP
300 W

Same across all providers.

Market comparison

$0.17/hr
Our median / hr
$0.32/hr
gpus.io median / hr
6
Our providers
2
gpus.io providers

View source on gpus.io

All-time price history

Offerings

All ×N×1×2×4×8In stock only · Partners only
ProviderCloudCountvCPURAMRegionPer GPU hrSPOTTotal/hr
Vast.ai
0.0% 30dReferral link
×1KR$0.12/hr$0.12/hrLaunch ↗
Vast.ai
0.0% 30dReferral link
×1KR$0.12/hr$0.12/hrLaunch ↗
HardwareHQ
0.0% 30dReferral link
×131 GBKR$0.14/hr$0.14/hrLaunch ↗
HardwareHQ
0.0% 30dReferral link
×124 GBVN$0.20/hr$0.20/hrLaunch ↗
HardwareHQ
0.0% 30dReferral link
×2110 GBCN$0.16/hr$0.32/hrLaunch ↗
HardwareHQ
0.0% 30dReferral link
×4220 GBCN$0.16/hr$0.64/hrLaunch ↗

LLMs that fit

ModelParams16-bit8-bit4-bit
Laguna-S 2.1
Poolside
118B 8B activeNeeds >8 GPUsNeeds >8 GPUsBarely fits on 7×
gpt-oss-120b
OpenAI
117B 5.1B activeNot releasedNot releasedBarely fits on 6×
Llama 3.1 70B
Meta
70BNeeds >8 GPUsBarely fits on 6×4× · 10 GB free
Qwen3.6 35B-A3B
Alibaba
35B 3B activeBarely fits on 7×Barely fits on 4×3× · 14.4 GB free
Qwen3 32B
Alibaba
32.8BBarely fits on 6×4× · 12.4 GB freeBarely fits on 2×
Qwen2.5 14B
Alibaba
14B3× · 12.6 GB free2× · 12.7 GB free1× · 4.6 GB free
Llama 3.1 8B
Meta
8B2× · 11.1 GB free1× · 4.6 GB free1× · 9.3 GB free
Mistral 7B
Mistral AI
7B2× · 13.8 GB free1× · 6 GB free1× · 10.1 GB free
Qwen3 4B
Alibaba
4B1× · 4.4 GB free1× · 9.6 GB free1× · 12.3 GB free

Estimates: model weights plus 20% for KV cache and runtime overhead, at short context lengths. Long contexts and large batches need more. Rows marked “tight” hold the weights but not that full margin. A ×N figure is the total VRAM across N of these GPUs and makes no claim about interconnect throughput.

Similar GPUs

Frequently asked questions

How much does the RTX 5070 Ti cost per hour?

The cheapest RTX 5070 Ti offering starts at $0.12/hr, and the market median is $0.17/hr.

Which cloud providers offer the RTX 5070 Ti?

We track RTX 5070 Ti offerings from Vast.ai, HardwareHQ.

How much VRAM does the RTX 5070 Ti have?

The RTX 5070 Ti has 16 GB of VRAM.

What LLMs can run on the RTX 5070 Ti?

With 16 GB of VRAM, the RTX 5070 Ti can run 9 of our tracked models in 4-bit quantization, including Laguna-S 2.1, gpt-oss-120b, Qwen3.6 35B-A3B. 6 models fit in 16-bit precision.

What's the biggest LLM I can run on multiple RTX 5070 Tis?

With 8× RTX 5070 Ti (128 GB total), the largest tracked model that fits in 16-bit is Qwen2.5 14B (14B params), which needs 3× GPUs.

Where is the RTX 5070 Ti cheapest?

The cheapest RTX 5070 Ti offering in our catalog is Vast.ai at $0.12/hr per GPU-hour (×1).