Back to all offerings

NVIDIA GTX 1650 Ti

4 GB VRAM55 W TDP0 providers0 offerings0 regions

The NVIDIA GTX 1650 Ti is a consumer-grade GPU built on the Turing architecture, featuring 4 GB of memory with 192 GB/s of memory bandwidth and a 55 W TDP. It delivers 3.8 TFLOPS of FP32 performance.

Cheapest
No current offerings
Median
across 0 offerings
Most expensive
90-day trend
-65.0%
median rate

Hardware specifications

Same across all providers

TDP
55 W
Memory bandwidth
192 GB/s
Architecture
Turing
FP32 performance
3.8 TFLOPS
Brand
NVIDIA
Series
GeForce GTX
Market price history

Median price across all providers

Prices updated 1 minute ago

0 offerings from 0 providers

Sorted by price ascending. Compare configurations side-by-side.

ProviderCountvCPURAMRegionPer GPU hourTotal/hrAction
Try broadening your filters or clear all filters to start over.

LLMs that fit on the NVIDIA GTX 1650 Ti

Estimated VRAM for popular open-weight LLMs at 16-bit (FP16/BF16), 8-bit (FP8/INT8) and 4-bit (INT4/MXFP4) precision, against 4 GB per GPU.

ModelParameters16-bit8-bit4-bit
Qwen3 4B
Alibaba
4B
9.7 GB
Too large
5.3 GB
Too large
3.1 GB
Barely fits on 1×

Estimates: model weights plus 20% for KV cache and runtime overhead, at short context lengths. Long contexts and large batches need more. Rows marked “tight” hold the weights but not that full margin. A ×N figure is the total VRAM across N of these GPUs and makes no claim about interconnect throughput. How we estimate this · Start from a model instead

Want more from this page?

See incorrect data?

Frequently asked questions

How much VRAM does the GTX 1650 Ti have?
The GTX 1650 Ti comes with 4 GB of VRAM, which determines the largest models and batch sizes it can hold in memory for training and inference.
Which LLMs can I run on the GTX 1650 Ti?
With 4 GB of VRAM, a single GTX 1650 Ti can serve models such as Qwen3 4B. The table above lists the estimated VRAM for each model at 16-bit, 8-bit and 4-bit precision, including the multi-GPU configurations needed for larger models.