95 devices · 4 brands

GPU database

Every GPU, laptop chip and Mac we track, with the numbers that decide local LLM performance: VRAM sets which models fit, memory bandwidth sets how fast they generate.

Specs verified 2026-10-07 · Prices verified 2026-09-29

Tracked954 brands
Most memory512 GBApple M5 Ultra
Fastest memory3,350 GB/sNVIDIA H100 80GB (SXM)
Cheapest 24 GB+$949Intel Arc Pro B70
Recommendation of the monthSeptember 2026

NVIDIA GeForce RTX 3090 24 GB (used)

At $600–750 used it is still the cheapest way to 24 GB of VRAM. It runs the same 27–32B models as an RTX 4090 at Q4 for about half the used price ($1,200–1,500), with 936 GB/s of memory bandwidth against the RTX 4090's 1,008 GB/s. Check VRAM health and thermals before you pay.

VRAM
24 GB
Bandwidth
936 GB/s
Fair used price
$600–750
Models that fit
124
DeviceVRAMBandwidth8B at Q4FitPrice
Apple M3 UltraApple Silicon · ARM, 3nm TSMC · 2025512 GB unified819 GB/s43–90 tok/s158—
Apple M5 UltraApple Silicon · ARM, 3nm TSMC · 2026512 GB unified1,229 GB/s59–123 tok/s158—
Apple M2 UltraApple Silicon · ARM, 5nm TSMC · 2023192 GB unified800 GB/s43–89 tok/s147—
Apple M1 UltraApple Silicon · ARM, 5nm TSMC · 2022128 GB unified800 GB/s43–89 tok/s143—
Apple M3 MaxApple Silicon · ARM, 3nm TSMC · 2023128 GB unified400 GB/s24–49 tok/s143—
Apple M4 MaxApple Silicon · ARM, 3nm TSMC · 2024128 GB unified546 GB/s31–64 tok/s143—
Apple M5 MaxApple Silicon · ARM, 3nm TSMC · 2026128 GB unified614 GB/s34–71 tok/s143—
Apple M2 MaxApple Silicon · ARM, 5nm TSMC · 202396 GB unified400 GB/s24–49 tok/s139—
Apple M3 Max (30-core GPU)Apple Silicon · ARM, 3nm TSMC · 202396 GB unified300 GB/s18–38 tok/s139—
NVIDIA RTX PRO 6000 BlackwellWorkstation GPU · Blackwell GB202 · 202596 GB1,792 GB/s113–234 tok/s143$8,565
NVIDIA RTX PRO 6000 Blackwell (Max-Q Edition)Workstation GPU · Blackwell GB202 · 2025-03-1896 GB1,792 GB/s113–234 tok/s143—
NVIDIA A100 80GB (PCIe)Datacenter GPU · Ampere GA100 · 202180 GB1,935 GB/s102–196 tok/s143$15,000
NVIDIA H100 80GB (PCIe)Datacenter GPU · Hopper GH100 · 202280 GB2,000 GB/s120–250 tok/s143$25,000
NVIDIA H100 80GB (SXM)Datacenter GPU · Hopper GH100 · 2022-09-2080 GB3,350 GB/s157–327 tok/s143—
NVIDIA RTX PRO 5000 Blackwell (72 GB)Workstation GPU · Blackwell GB202 · 2025-03-1872 GB1,344 GB/s94–194 tok/s139—
Apple M1 MaxApple Silicon · ARM, 5nm TSMC · 202164 GB unified400 GB/s24–49 tok/s131—
Apple M5 ProApple Silicon · ARM, 3nm TSMC · 202664 GB unified307 GB/s19–39 tok/s131—
NVIDIA L40Datacenter GPU · Ada Lovelace AD102 · 2022-10-1348 GB864 GB/s77–147 tok/s131—
NVIDIA L40SDatacenter GPU · Ada Lovelace AD102 · 202348 GB864 GB/s77–147 tok/s131$7,499
AMD Radeon PRO W7900Workstation GPU · RDNA 3 Navi 31 · 2023-04-1348 GB864 GB/s50–103 tok/s131—
NVIDIA RTX 6000 Ada GenerationWorkstation GPU · Ada Lovelace AD102 · 202248 GB960 GB/s83–159 tok/s131$6,799
NVIDIA RTX A6000Workstation GPU · Ampere GA102 · 2020-10-05used48 GB768 GB/s51–98 tok/s131$4,649
NVIDIA RTX PRO 5000 Blackwell (48 GB)Workstation GPU · Blackwell GB202 · 2025-03-1848 GB1,344 GB/s94–194 tok/s131—
Apple M3 ProApple Silicon · ARM, 3nm TSMC · 202336 GB unified153 GB/s9.7–20 tok/s124—
Apple M4 Max (32-core GPU)Apple Silicon · ARM, 3nm TSMC · 202436 GB unified410 GB/s24–50 tok/s124—
Apple M5 Max (32-core GPU)Apple Silicon · ARM, 3nm TSMC · 202636 GB unified460 GB/s27–56 tok/s124—
AMD Instinct MI50 32GBDatacenter GPU · GCN 5 Vega 20 · 2018-11-18used32 GB1,024 GB/s—124—
Apple M1 ProApple Silicon · ARM, 5nm TSMC · 202132 GB unified200 GB/s13–26 tok/s124—
Apple M2 ProApple Silicon · ARM, 5nm TSMC · 202332 GB unified200 GB/s13–26 tok/s124—
Apple M4Apple Silicon · ARM, 3nm TSMC · 202432 GB unified120 GB/s7.7–16 tok/s124—
Apple M5Apple Silicon · ARM, 3nm TSMC · 202532 GB unified153 GB/s9.7–20 tok/s124—
Apple M6Apple Silicon · ARM, 2nm TSMC · 202632 GB unified170 GB/s11–22 tok/s124—
Intel Arc Pro B65Workstation GPU · Xe2 Battlemage BMG-G31 · 202632 GB608 GB/s35–72 tok/s124—
Intel Arc Pro B70Workstation GPU · Xe2 Battlemage BMG-G31 · 202632 GB608 GB/s35–72 tok/s124$949
AMD Radeon AI PRO R9700Workstation GPU · RDNA 4 Navi 48 · 202532 GB640 GB/s42–87 tok/s124$1,299
NVIDIA RTX 5000 Ada GenerationWorkstation GPU · Ada Lovelace AD102 · 2023-08-0932 GB576 GB/s56–107 tok/s124—
NVIDIA GeForce RTX 5090Consumer GPU · Blackwell GB202 · 202532 GB1,792 GB/s113–234 tok/s124$4,700
NVIDIA RTX PRO 4500 Blackwell (Workstation Edition)Workstation GPU · Blackwell GB203 · 2025-03-1832 GB896 GB/s70–145 tok/s124—
Apple M2Apple Silicon · ARM, 5nm TSMC · 202224 GB unified100 GB/s6.4–13 tok/s95—
Apple M3Apple Silicon · ARM, 3nm TSMC · 202324 GB unified100 GB/s6.4–13 tok/s95—
Apple M4 ProApple Silicon · ARM, 3nm TSMC · 202424 GB unified273 GB/s17–35 tok/s95—
Intel Arc Pro B60Workstation GPU · Xe2 Battlemage BMG-G21 · 202524 GB456 GB/s28–57 tok/s124—
NVIDIA A10Datacenter GPU · Ampere GA102 · 2021-04-1224 GB600 GB/s41–79 tok/s124—
NVIDIA L4Datacenter GPU · Ada Lovelace AD104 · 2023-03-2124 GB300 GB/s32–61 tok/s124—
NVIDIA GeForce RTX 3090Consumer GPU · Ampere GA102 · 202024 GB936 GB/s60–115 tok/s124$1,499
NVIDIA GeForce RTX 3090 TiLegacy / used GPU · Ampere GA102 · 2022used24 GB1,008 GB/s63–122 tok/s124$1,999
NVIDIA GeForce RTX 4090Consumer GPU · Ada Lovelace AD102 · 202224 GB1,008 GB/s86–165 tok/s124$1,599
NVIDIA RTX A5000Workstation GPU · Ampere GA102 · 2021-04-12used24 GB768 GB/s51–98 tok/s124—
NVIDIA RTX PRO 4000 BlackwellWorkstation GPU · Blackwell GB203 · 2025-03-1824 GB672 GB/s56–116 tok/s124—
AMD Radeon RX 7900 XTXConsumer GPU · RDNA 3 Navi 31 · 202224 GB960 GB/s54–112 tok/s124$999
NVIDIA Tesla P40Datacenter GPU · Pascal GP102 · 2016-09-13used24 GB346 GB/s—124—
AMD Radeon RX 7900 XTConsumer GPU · RDNA 3 Navi 31 · 202220 GB800 GB/s47–98 tok/s106$899
Apple M1Apple Silicon · ARM, 5nm TSMC · 202016 GB unified68 GB/s4.4–9.2 tok/s77—
Intel Arc Pro B50Workstation GPU · Xe2 Battlemage BMG-G21 · 202516 GB224 GB/s15–31 tok/s87$349
NVIDIA GeForce RTX 4060 Ti 16GBConsumer GPU · Ada Lovelace AD106 · 202316 GB288 GB/s31–59 tok/s87$499
NVIDIA GeForce RTX 4070 Ti SuperConsumer GPU · Ada Lovelace AD103 · 202416 GB672 GB/s63–121 tok/s87$799
NVIDIA GeForce RTX 4080Consumer GPU · Ada Lovelace AD103 · 202216 GB717 GB/s66–127 tok/s87$1,199
NVIDIA GeForce RTX 4080 SuperConsumer GPU · Ada Lovelace AD103 · 202416 GB736 GB/s68–130 tok/s87$999
NVIDIA GeForce RTX 5060 Ti 16GBConsumer GPU · Blackwell GB206 · 202516 GB448 GB/s40–83 tok/s87$805
NVIDIA GeForce RTX 5070 TiConsumer GPU · Blackwell GB205 · 202516 GB896 GB/s70–145 tok/s87$749
NVIDIA GeForce RTX 5080Consumer GPU · Blackwell GB203 · 202516 GB960 GB/s74–153 tok/s87$999
NVIDIA RTX PRO 2000 BlackwellWorkstation GPU · Blackwell GB206 · 2025-08-1116 GB288 GB/s27–56 tok/s87—
AMD Radeon RX 6800 XTLegacy / used GPU · RDNA 2 Navi 21 · 2020-11-18used16 GB512 GB/s—87—
AMD Radeon RX 6900 XTLegacy / used GPU · RDNA 2 Navi 21 · 2020-12-08used16 GB512 GB/s—87—
AMD Radeon RX 7600 XTConsumer GPU · RDNA 3 Navi 33 · 2024-01-2416 GB288 GB/s20–42 tok/s87—
AMD Radeon RX 7800 XTConsumer GPU · RDNA 3 Navi 32 · 202316 GB624 GB/s39–81 tok/s87$499
AMD Radeon RX 7900 GREConsumer GPU · RDNA 3 Navi 31 · 2023-07-2716 GB576 GB/s36–76 tok/s87—
AMD Radeon RX 9060 XT 16GBConsumer GPU · RDNA 4 Navi 44 · 202516 GB320 GB/s24–49 tok/s87$349
AMD Radeon RX 9070Consumer GPU · RDNA 4 Navi 48 · 202516 GB640 GB/s42–87 tok/s87$549
AMD Radeon RX 9070 XTConsumer GPU · RDNA 4 Navi 48 · 202516 GB640 GB/s42–87 tok/s87$599
Intel Arc B580Consumer GPU · Xe2 Battlemage BMG-G21 · 202412 GB456 GB/s28–57 tok/s77$249
NVIDIA GeForce RTX 3060 (12GB)Consumer GPU · Ampere GA106 · 202112 GB360 GB/s26–50 tok/s77$329
NVIDIA GeForce RTX 3080 12GBLegacy / used GPU · Ampere GA102 · 2022used12 GB912 GB/s59–112 tok/s77—
NVIDIA GeForce RTX 3080 TiLegacy / used GPU · Ampere GA102 · 2021used12 GB912 GB/s59–112 tok/s77$1,199
NVIDIA GeForce RTX 4070Consumer GPU · Ada Lovelace AD104 · 202312 GB504 GB/s50–96 tok/s77$599
NVIDIA GeForce RTX 4070 SuperConsumer GPU · Ada Lovelace AD104 · 202412 GB504 GB/s50–96 tok/s77$599
NVIDIA GeForce RTX 4070 TiConsumer GPU · Ada Lovelace AD104 · 202312 GB504 GB/s50–96 tok/s77$799
NVIDIA GeForce RTX 5070Consumer GPU · Blackwell GB205 · 202512 GB672 GB/s56–116 tok/s77$549
AMD Radeon RX 6700 XTLegacy / used GPU · RDNA 2 Navi 22 · 2021-03-01used12 GB384 GB/s—77—
AMD Radeon RX 7700 XTConsumer GPU · RDNA 3 Navi 32 · 202312 GB432 GB/s29–60 tok/s77$449
AMD Radeon RX 9070 GREConsumer GPU · RDNA 4 Navi 48 · 2025-05-0812 GB432 GB/s31–64 tok/s77—
NVIDIA GeForce GTX 1080 TiLegacy / used GPU · Pascal GP102 · 2017-03-10used11 GB484 GB/s—77$699
NVIDIA GeForce RTX 2080 TiLegacy / used GPU · Turing TU102 · 2018-09-27used11 GB616 GB/s—77$999
Intel Arc B570Consumer GPU · Xe2 Battlemage BMG-G21 · 202510 GB380 GB/s24–49 tok/s75$219
NVIDIA GeForce RTX 3080 (10GB)Consumer GPU · Ampere GA102 · 202010 GB760 GB/s50–97 tok/s75$699
NVIDIA GeForce RTX 3060 TiLegacy / used GPU · Ampere GA104 · 2020-12-02used8 GB448 GB/s32–61 tok/s65$399
NVIDIA GeForce RTX 3070Consumer GPU · Ampere GA104 · 20208 GB448 GB/s32–61 tok/s65$499
NVIDIA GeForce RTX 3070 TiConsumer GPU · Ampere GA104 · 20218 GB608 GB/s42–80 tok/s65$599
NVIDIA GeForce RTX 4060Consumer GPU · Ada Lovelace AD107 · 20238 GB272 GB/s29–56 tok/s65$299
NVIDIA GeForce RTX 4060 Ti 8GBConsumer GPU · Ada Lovelace AD106 · 20238 GB288 GB/s31–59 tok/s65$399
NVIDIA GeForce RTX 5050Consumer GPU · Blackwell GB207 · 2025-07-018 GB320 GB/s30–61 tok/s65$249
NVIDIA GeForce RTX 5060Consumer GPU · Blackwell GB206 · 20258 GB448 GB/s40–83 tok/s65$299
NVIDIA GeForce RTX 5060 Ti 8GBConsumer GPU · Blackwell GB206 · 20258 GB448 GB/s40–83 tok/s65$379
AMD Radeon RX 7600Consumer GPU · RDNA 3 Navi 33 · 2023-05-248 GB288 GB/s20–42 tok/s65—
AMD Radeon RX 9060 XT 8GBConsumer GPU · RDNA 4 Navi 44 · 20258 GB320 GB/s24–49 tok/s65$299

Quick buying guide

Under $300
RTX 3060 12GB12 GB VRAM, runs all 7–8B models
Under $500
RTX 4060 Ti 16GBbest budget 16 GB card
$1,000–1,600
RTX 4090 24GBfastest single consumer card; 32B-class models at Q4_K_M
Mac laptop
M3 Pro 36GB or M4 Prosilent, power-efficient
Mac desktop
M4 Max 128GB~96 GB usable: 70B-class models at Q4_K_M

Speed is Llama 3.1 8B at Q4_K_M: a bandwidth-roofline estimate shown as a range unless the device page marks it measured. Mac memory is the full unified pool; about 75% of it is usable by a model. Fit counts catalogued model variants whose Q4_K_M weights plus runtime overhead fit the usable memory, before any context is added. Methodology