GPU comparison · updated

Pick the right GPU for local AI — price vs. performance.

Compare NVIDIA, AMD, Intel, and Apple Silicon on memory bandwidth, VRAM, and bandwidth-per-euro value. Real specs, real prices — no marketing numbers.

GPUs
Brands
VRAM range
Peak mem BW
Live recommendation
Min VRAM
Loading data…
01 /Comparison

Filter by model size to see which GPUs can run it, then sort by what matters: throughput, value, or VRAM.

Price vs VRAM
Upper-left = great value. Lower-right = avoid. Click a point to open the GPU detail page.
Y axis
NVIDIA AMD Intel Apple
Brand
Tier
Min VRAM
results
GPU VRAM Mem BW TDP Street BW/€
Loading data…
02 /Picks for this model

Three angles — fastest bandwidth under budget, best bandwidth per €, and most VRAM.

03 /Methodology

Bandwidth-derived speed estimates from hardware specs. Real street prices where available, MSRP fallback otherwise.

01

Bandwidth as speed proxy

Token generation at batch size 1 is memory-bandwidth-bound: every forward pass streams the full model weights from VRAM. Higher GB/s means faster generation, regardless of compute units. Bandwidth is therefore the most honest single-number proxy for inference speed without model-specific assumptions.

02

Hardware specs

VRAM, memory bandwidth, TDP, and clock data from official manufacturer specs and verified press releases.

03

Street prices

Real market prices from Geizhals.de where available. MSRP used as fallback for workstation and Apple Silicon.

04

Relative comparison

All figures are from manufacturer specs. Use them to compare GPUs against each other — not as absolute performance targets. Real-world inference speed depends on driver version, software stack, model quantisation, and context length.

0 selected