Homelab & AI

AI Workload GPU Price Tracker

Consumer RTX, Workstation RTX A-Series & Datacenter Tesla / A100 / Instinct

For running local LLMs, Stable Diffusion, or any other AI workload, VRAM, not raw gaming benchmarks, is usually the number that decides what you can load. This tracks current $/GB VRAM pricing across the GPUs homelabbers actually buy for AI: consumer RTX 3090/4090 cards, workstation RTX A-series cards, and used datacenter compute cards like the Tesla P40/V100, NVIDIA L4/A100, and AMD Instinct MI50.

This is a periodically hand-checked catalog rather than a continuously live feed — see the All Trackers FAQ for how each tracker's freshness works.

Price Alert
Never miss a price drop
Email me when it drops to
$
/GB VRAM or lower

One click to confirm, one click to unsubscribe — no account needed.

Why VRAM Is the Number That Matters

A GPU's shader count and memory bandwidth affect how fast a model runs, but VRAM decides whether the model fits at all. If a model doesn't fit in VRAM, it either won't load or has to spill over to much slower system RAM or disk offloading. That makes $/GB VRAM, not just raw price, the useful number to compare across wildly different card types.

This is also why older datacenter cards like the Tesla P40 stay popular in homelab AI circles years after their gaming-era equivalents would be considered obsolete: 24GB of VRAM still fits plenty of useful quantized models, and depreciated enterprise surplus pricing means that VRAM is often the cheapest on the used market of any card type here.

Quantization (running a model at reduced numeric precision, like 4-bit instead of 16-bit) is the other lever. It roughly quarters the VRAM a given model needs at some cost to output quality, which is why a 24GB card can run models that would need 80GB+ at full precision.

The Three GPU Classes on This Page

Consumer (RTX 3090 / 4090)

Usually the best raw price per GB of VRAM, since they're mass-produced gaming cards. No ECC memory, and warranty/duty-cycle expectations assume gaming use, not 24/7 compute.

Prosumer (RTX A-Series)

ECC memory, blower-style coolers built for dense multi-GPU chassis, and higher VRAM configurations than their consumer counterparts, at a real price premium.

Datacenter (Tesla / A100 / Instinct)

Purpose-built for sustained compute, often the cheapest $/GB VRAM on older generations as enterprise fleets retire them. Expect passive cooling that needs a server chassis or an aftermarket fan, plus a 250W+ draw worth running through my power cost calculator if it'll run 24/7.

Clear filters

26 GPUs found · Consumer · prices checked as of Oct 11, 2026

GPUBrandClassConditionVRAMPrice$/GB VRAMDeal

Bracket For ZOTAC GeForce RTX 3060Ti RTX3070 RTX3080 RTX3090 Graphics Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$7.58$0.32/GB

Baffle Plate Bracket for EVGA RTX 3060ti 3070 RTX 3080 3090 FTW3 XC3 Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$11.50$0.48/GB

Baffle Plate Bracket for Zotac RTX 3060 RTX 3070 RTX 3080 RTX 3090 Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$14.88$0.62/GB

Baffle Plate Bracket for Zotac RTX 3060 RTX 3080 RTX 3070 RTX 3090 Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$14.99$0.62/GB

Baffle Plate Bracket for EVGA RTX 3060ti 3070 RTX 3080 3090 FTW3 XC3 Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$15.00$0.63/GB

Baffle Plate Bracket for Zotac RTX 3060 RTX 3070 RTX 3080 RTX 3090 Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$15.00$0.63/GB

Baffle Plate Bracket for Zotac RTX 3060 RTX 3070 RTX 3080 RTX 3090 Video Card

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$39.00$1.63/GB

NVIDIA GEFORCE 30 SERIES RTX 3070 3080 3090 Ti VIDEO GRAPHICS CARD GPU REPAIR

RTX 3090 Ti · eBay

NVIDIAConsumerUsed24 GB$39.95$1.66/GB

heatsink for RTX3080 3080ti 3090

RTX 3090 Ti · eBay

NVIDIAConsumerUsed24 GB$88.00$3.67/GB

heat sink for colorful3090Ti (Removed GPU and RAM )

RTX 3090 Ti · eBay

NVIDIAConsumerUsed24 GB$99.00$4.13/GB

PNY NVIDIA GeForce RTX 4090 24GB GDDR6X COOLER ONLY NO GPU - Heatsink only

RTX 4090 · eBay

NVIDIAConsumerUsed24 GB$118.01$4.92/GB

HP RTX 4090 24GB Graphics Card No Core No Memory

RTX 4090 · eBay

NVIDIAConsumerUsed24 GB$149.20$6.22/GB

CoolerAge Graphics Card Heatsink for Palit RTX 3080 Ti 3090 Gamingpro 87mm

RTX 3090 Ti · eBay

NVIDIAConsumerNew24 GB$195.33$8.14/GB

Water cooler for ASUS ROG Strix LC GeForce RTX 4090 OC 24GB GDDR6X Graphics Card

RTX 4090 · eBay

NVIDIAConsumerNew24 GB$199.00$8.29/GB

Water cooler for ASUS ROG Strix LC GeForce RTX 4090 OC 24GB GDDR6X Graphics Card

RTX 4090 · eBay

NVIDIAConsumerUsed24 GB$289.99$12.08/GB

2X ASUS TUF GeForce RTX 4090 24GB OG - TUF-RTX4090-24G-OG-GAMING ENCLOSURE ONLY

RTX 4090 · eBay

NVIDIAConsumerUsed24 GB$300.00$12.50/GB

3070m-8G GPU Graphics Card

RTX 3090 · eBay

NVIDIAConsumerUsed24 GB$418.00$17.42/GB

DELL NVIDIA GeForce RTX 3060 3080 3070 3090 GDDR6X 8/10/12/24GB GRAPHICS CARD

RTX 3090 · eBay

NVIDIAConsumerUsed24 GB$494.76$20.62/GB

DELL NVIDIA GeForce RTX 3060/3070/3080/3090 8/10/12/24GB GDDR6X GRAPHICS CARD

RTX 3090 · eBay

NVIDIAConsumerUsed24 GB$546.56$22.77/GB

ZOTAC GAMING GeForce RTX 3090 Trinity OC 24GB Graphics Card DOES NOT BOOT READ

RTX 3090 · eBay

NVIDIAConsumerUsed24 GB$549.99$22.92/GB

Dell NVIDIA GeForce RTX 3060 3070 3080 3090 8 10 12 24GB GDDR6X GRAPHICS CARD

RTX 3090 · eBay

NVIDIAConsumerNew24 GB$563.73$23.49/GB

Dell NVIDIA GeForce RTX 3060 3070 3080 3090 8 10 12 24GB GDDR6X GRAPHICS CARD

RTX 3090 · eBay

NVIDIAConsumerNew24 GB$563.73$23.49/GB

Dell NVIDIA GeForce RTX 3060 3070 3080 3090 8 10 12 24GB GDDR6X GRAPHICS CARD

RTX 3090 · eBay

NVIDIAConsumerNew24 GB$563.73$23.49/GB

Colorful GeForce GTX 1070 8GB GDDR5 Dual Fan Gaming Graphics Card

RTX 3090 · eBay

NVIDIAConsumerUsed8 GB$249.00$31.13/GB

Broken Asus RTX 3090 TUF Gaming 24GB GDDR6X Graphics Card TUF-RTX3090-024G-2I3S

RTX 3090 · eBay

NVIDIAConsumerUsed24 GB$749.99$31.25/GB

VIDIA RTX 4060 Ti Founders Edition 8GB GDDR6 ITX SFF Small Form Factor

RTX 3090 · eBay

NVIDIAConsumerUsed8 GB$748.00$93.50/GB
Price history: cheapest $/GB VRAM over time

The lowest $/GB VRAM found each day this tracker has logged live listings — useful for seeing whether now is a cheap or expensive time to buy, not just what's cheapest today.

Loading price history…

Frequently Asked Questions

How much VRAM do I need to run a local LLM?

As a rough rule of thumb, a model at full FP16 precision needs about 2GB of VRAM per billion parameters, while a 4-bit quantized version needs roughly 0.5-0.7GB per billion parameters. That puts a quantized 7-8B model at 8-12GB, a 13B model at around 16GB, a 30-34B model at 24GB, and a 70B model at 48GB or more. Larger frontier-scale models generally require 80GB+ or splitting the model across multiple GPUs. If you're sizing a card specifically to run a self-hosted AI agent like OpenClaw, see how I sized mine on a Proxmox homelab.

Do Tesla and other datacenter GPUs need extra cooling to use outside a server?

In most cases, yes. Cards like the Tesla P40, P100, and V100 (PCIe versions) are passively cooled and rely on the high static-pressure airflow inside a server chassis to stay cool. Running one in a desktop case or open-air rig typically requires an aftermarket blower fan or 3D-printed shroud, or it will overheat and throttle.

Is a used mining GPU safe to buy for AI compute?

For the most part, yes. Mining is a steady, moderate 24/7 load that's arguably gentler on a card than the thermal cycling of gaming, and the VRAM chips that matter most for AI work see minimal wear either way. Watch listings for mentions of reflowed or reballed memory, and buy from sellers with a return policy so you can verify the card before the window closes.

What's the difference between consumer, prosumer, and datacenter GPUs for AI?

Consumer cards (RTX 3090/4090) offer the best raw price per GB of VRAM but are built for gaming, without ECC memory or long-duty-cycle warranty terms. Prosumer/workstation cards (RTX A-series) add ECC memory, blower coolers suited to dense multi-GPU builds, and higher VRAM configurations at a price premium. Datacenter cards (Tesla, A100, L4, Instinct) are purpose-built for sustained compute and often have the lowest $/GB VRAM on older generations, but need server-style airflow and sometimes different power connectors.

Can I combine multiple GPUs to run a bigger model?

Yes. Tools like llama.cpp, vLLM, and Ollama support splitting a model's layers across multiple GPUs' VRAM. Most post-30-series consumer cards no longer support NVLink, so multi-GPU setups communicate over PCIe instead, which is slower than NVLink but works fine for inference. Prosumer and datacenter cards more often retain NVLink or NVSwitch support for tighter multi-GPU scaling.

Disclaimer: This is a personal tracking tool I built for my own use, not a professional pricing or comparison service. Listings, prices, and outbound links are pulled from third-party sources and can be stale, inaccurate, or broken. I'm not responsible for their accuracy or where a link leads. Always verify details directly with the retailer or source before you rely on them.