Back to all offerings

NVIDIA GTX 1660

6 GB VRAM120 W TDP1 providers1 offerings1 regions

The NVIDIA GTX 1660 is a consumer-grade GPU built on the Turing architecture, featuring 6 GB of memory with 192 GB/s of memory bandwidth and a 120 W TDP. It delivers 5 TFLOPS of FP32 performance. Currently available from 1 provider starting at $0.68/GPU/hour, with a market median of $0.68/GPU/hour across 1 configurations.

Cheapest
$0.680/hr
Vast.ai
Median
$0.680/hr
across 1 offerings
Most expensive
$0.680/hr
Vast.ai
90-day trend
+580.0%
median rate

Hardware specifications

Same across all providers

TDP
120 W
Memory bandwidth
192 GB/s
Architecture
Turing
FP32 performance
5 TFLOPS
Brand
NVIDIA
Series
GeForce GTX
Market price history

Median price across all providers

Prices updated 20 minutes ago

1 offering from 1 providers

Sorted by price ascending. Compare configurations side-by-side.

ProviderCountvCPURAMRegionPer GPU hourTotal/hrAction
Vast.ai logoVast.aiReferral linkTrendingCheapest
×148 GB
Vietnam
$0.680
-44.9% 30d
$0.68
Launch

LLMs that fit on the NVIDIA GTX 1660

Estimated VRAM for popular open-weight LLMs at 16-bit (FP16/BF16), 8-bit (FP8/INT8) and 4-bit (INT4/MXFP4) precision, against 6 GB per GPU.

ModelParameters16-bit8-bit4-bit
Granite 4.1 8B
IBM
8.8B
21 GB
Too large
12 GB
Too large
6.8 GB
Tight on 1× · 6% spare
Qwen3 8B
Alibaba
8.2B
20 GB
Too large
11 GB
Too large
6.3 GB
Tight on 1× · 14% spare
Llama 3.1 8B Instruct
Meta
8B
19 GB
Too large
11 GB
Too large
6.2 GB
Tight on 1× · 16% spare
Mistral 7B Instruct v0.3
Mistral AI
7.2B
17 GB
Too large
9.6 GB
Too large
5.6 GB
Barely fits on 1×
Qwen3 4B
Alibaba
4B
9.7 GB
Too large
5.3 GB
Barely fits on 1×
3.1 GB
1× · 2.9 GB free

Estimates: model weights plus 20% for KV cache and runtime overhead, at short context lengths. Long contexts and large batches need more. Rows marked “tight” hold the weights but not that full margin. A ×N figure is the total VRAM across N of these GPUs and makes no claim about interconnect throughput. How we estimate this · Start from a model instead

Want more from this page?

See incorrect data?

Frequently asked questions

How much does the GTX 1660 cost per hour?
The GTX 1660 starts at $0.68 per GPU-hour, with a median of $0.68 across 1 available configurations. Actual pricing depends on the provider, region, and whether you rent on-demand, spot or reserved.
Which cloud providers offer the GTX 1660?
1 providers currently list the GTX 1660: Vast.ai. Compare their hourly pricing, regions and machine specs in the table above.
How much VRAM does the GTX 1660 have?
The GTX 1660 comes with 6 GB of VRAM, which determines the largest models and batch sizes it can hold in memory for training and inference.
Which LLMs can I run on the GTX 1660?
With 6 GB of VRAM, a single GTX 1660 can serve models such as Mistral 7B Instruct v0.3, Qwen3 4B. The table above lists the estimated VRAM for each model at 16-bit, 8-bit and 4-bit precision, including the multi-GPU configurations needed for larger models.
Where is the GTX 1660 cheapest?
The lowest on-demand price we currently track for the GTX 1660 is $0.68 per GPU-hour at Vast.ai. Because providers update pricing and availability frequently, check the live table above before you launch.