
Choosing a GPU for Running LLMs at Home
VRAM is king, but it is not the whole story. A practical guide to picking a card for local inference and light fine-tuning.
Diego Ramos🇧🇷 Value & Buying CorrespondentJun 25, 2026 6m readIf you take one thing from this guide: buy the most VRAM you can afford.
Why VRAM dominates
Model size, context length, and batch size all consume memory. Run out, and performance collapses as data spills to system RAM. A card with more memory but slightly slower cores will beat a faster card that cannot fit your model.
- Entry: enough VRAM for quantized mid-size models
- Sweet spot: a card that fits popular open models comfortably
- Enthusiast: multi-card or high-memory workstation GPUs
Quantization stretches your VRAM, but it is not a substitute for having enough in the first place.
Beyond the card
Do not forget power supply headroom and case airflow. A starved or thermally throttled GPU quietly loses you performance.

🇧🇷 Value & Buying Correspondent · São Paulo, Brazil
Finds the smart buy — the best value for what you actually do.

Calculus I
by Richard Murdoch Montgomery
Limits, derivatives, integrals, and series — a first course in calculus with formal proofs, worked examples, and applications to physics and engineering.

Treatise on Systems Biology
by Richard Murdoch Montgomery
Modelling gene regulatory networks, metabolic pathways, and ecological dynamics — where mathematics meets molecular biology.

Glioblastoma Growth Modelling
by Richard Murdoch Montgomery
Mathematical oncology meets computational neuroscience — reaction-diffusion models, imaging-driven simulations, and treatment optimisation.

The Scientific Financial Calculator 12C: Finance
by Richard Murdoch Montgomery
Over 600 pages and 51 chapters on the HP 12C — bond pricing, duration, convexity, portfolio mathematics, and regression analysis.
Comments
Open discussion — no account needed. Be respectful.
More from Hardware Buying Guides
Intel Gaudi 3 vs AMD Instinct MI300X and MI325X: The Enterprise AI Accelerator Buying Guide for 2026
Enterprise AI accelerators have never been more capable — or more confusing to procure. This spec-grounded guide compares Intel Gaudi 3, AMD Instinct MI300X, and MI325X on HBM capacity, memory bandwidth, software ecosystems, and real deployment requirements, so you can make a defensible purchase decision.
Kaito TanakaBest USB4 and Thunderbolt 5 External Storage for AI/ML Dataset Management in 2026
USB4 40Gbps portable SSDs and Thunderbolt 5 enclosures have finally made external storage fast enough to matter for AI/ML workflows — here's how to pick the right drive for dataset ingest, model staging, and checkpoint backup without overspending.
Diego RamosApple Silicon for Local AI Inference in 2026: Mac mini M4 Pro vs Mac Studio M4 Max vs M3 Ultra
The Mac Pro M4 Ultra never shipped — but Apple's current Mac mini and Mac Studio lineup still offers the most compelling unified-memory inference platform money can buy. Here's exactly which configuration to choose, and when a discrete GPU rig beats them all.
Kaito Tanaka