Back to all offerings

NVIDIA RTX 2060

6 GB VRAM160 W TDP1 providers2 offerings2 regions

The NVIDIA RTX 2060 is a consumer-grade GPU built on the Turing architecture, featuring 6 GB of memory with 336 GB/s of memory bandwidth and a 160 W TDP. It delivers 6.5 TFLOPS of FP32 performance. Currently available from 1 provider starting at $0.07/GPU/hour, with a market median of $0.20/GPU/hour across 2 configurations.

Cheapest
$0.070/hr
Vast.ai
Median
$0.195/hr
across 2 offerings
Most expensive
$0.320/hr
Vast.ai
90-day trend
+290.0%
median rate

Hardware specifications

Same across all providers

TDP
160 W
Memory bandwidth
336 GB/s
Architecture
Turing
FP32 performance
6.5 TFLOPS
Brand
NVIDIA
Series
GeForce RTX
Market price history

Median price across all providers

Prices updated 2 hours ago

1 offering from 1 providers

Sorted by price ascending. Compare configurations side-by-side.

ProviderCountvCPURAMRegionPer GPU hourTotal/hrAction
Vast.ai logoVast.aiReferral linkTrendingCheapest
×112 GB
BhutanUnited States
From$0.070
-9.3% 30d
From$0.07
Launch

LLMs that fit on the NVIDIA RTX 2060

Estimated VRAM for popular open-weight LLMs at 16-bit (FP16/BF16), 8-bit (FP8/INT8) and 4-bit (INT4/MXFP4) precision, against 6 GB per GPU.

ModelParameters16-bit8-bit4-bit
Granite 4.1 8B
IBM
8.8B
21 GB
Too large
12 GB
Too large
6.8 GB
Tight on 1× · 6% spare
Qwen3 8B
Alibaba
8.2B
20 GB
Too large
11 GB
Too large
6.3 GB
Tight on 1× · 14% spare
Llama 3.1 8B Instruct
Meta
8B
19 GB
Too large
11 GB
Too large
6.2 GB
Tight on 1× · 16% spare
Mistral 7B Instruct v0.3
Mistral AI
7.2B
17 GB
Too large
9.6 GB
Too large
5.6 GB
Barely fits on 1×
Qwen3 4B
Alibaba
4B
9.7 GB
Too large
5.3 GB
Barely fits on 1×
3.1 GB
1× · 2.9 GB free

Estimates: model weights plus 20% for KV cache and runtime overhead, at short context lengths. Long contexts and large batches need more. Rows marked “tight” hold the weights but not that full margin. A ×N figure is the total VRAM across N of these GPUs and makes no claim about interconnect throughput. How we estimate this · Start from a model instead

Want more from this page?

See incorrect data?

Frequently asked questions

How much does the RTX 2060 cost per hour?
The RTX 2060 starts at $0.07 per GPU-hour, with a median of $0.20 across 2 available configurations. Actual pricing depends on the provider, region, and whether you rent on-demand, spot or reserved.
Which cloud providers offer the RTX 2060?
1 providers currently list the RTX 2060: Vast.ai. Compare their hourly pricing, regions and machine specs in the table above.
How much VRAM does the RTX 2060 have?
The RTX 2060 comes with 6 GB of VRAM, which determines the largest models and batch sizes it can hold in memory for training and inference.
Which LLMs can I run on the RTX 2060?
With 6 GB of VRAM, a single RTX 2060 can serve models such as Mistral 7B Instruct v0.3, Qwen3 4B. The table above lists the estimated VRAM for each model at 16-bit, 8-bit and 4-bit precision, including the multi-GPU configurations needed for larger models.
Where is the RTX 2060 cheapest?
The lowest on-demand price we currently track for the RTX 2060 is $0.07 per GPU-hour at Vast.ai. Because providers update pricing and availability frequently, check the live table above before you launch.