BLK · LEADERBOARD

RunLocalAI Score

Every catalog hardware unit ranked by composite score (0–1000): measured tok/s, VRAM fit, ecosystem support, perf-per-watt. 2 of 157 ranks anchored to a measured benchmark — the rest are honestly flagged as extrapolated or estimated.

Methodology: /methodology · Run your own: curl -fsSL runlocalai.co/bench.mjs -o bench.mjs && node bench.mjs

23 units shown · sorted by score

#HardwareTierScoreData
1NVIDIA H20 (96GB)
nvidia · workstation · 96GB
B697Estimated
2NVIDIA H200 NVL (PCIe)
nvidia · workstation · 141GB
B684Estimated
3NVIDIA B200
nvidia · workstation · 192GB
B684Estimated
4NVIDIA H200
nvidia · workstation · 141GB
B676Estimated
5NVIDIA B300 (Blackwell Ultra)
nvidia · workstation · 288GB
B669Estimated
6NVIDIA H100 NVL
nvidia · workstation · 188GB
B663Estimated
7NVIDIA H100 PCIe
nvidia · workstation · 80GB
B662Estimated
8NVIDIA A100 80GB SXM
nvidia · workstation · 80GB
B657Estimated
9NVIDIA H100 SXM
nvidia · workstation · 80GB
B655Estimated
10NVIDIA RTX PRO 6000 Blackwell
nvidia · workstation · 96GB
B650Estimated
11NVIDIA A100 40GB
nvidia · workstation · 40GB
B635Estimated
12NVIDIA GB200 NVL72
nvidia · workstation · 13824GB
B631Estimated
13NVIDIA GeForce RTX 5090
nvidia · enthusiast · 32GB
B630Estimated
14NVIDIA RTX 4090 48GB (China-mod)
nvidia · workstation · 48GB
B534Estimated
15NVIDIA RTX 5000 PRO Blackwell 48GB
nvidia · workstation · 48GB
B529Estimated
16NVIDIA RTX 6000 Ada Generation
nvidia · workstation · 48GB
B529Estimated
17NVIDIA GeForce RTX 3090 Ti
nvidia · enthusiast · 24GB
B520Estimated
18NVIDIA GeForce RTX 4090
nvidia · enthusiast · 24GB
B520Estimated
19NVIDIA GeForce RTX 5090 Mobile
nvidia · enthusiast · 24GB
B512Estimated
20NVIDIA RTX PRO 4500 Blackwell
nvidia · workstation · 32GB
B507Estimated
21NVIDIA GeForce RTX 3090
nvidia · enthusiast · 24GB
B505Estimated
22NVIDIA L40
nvidia · workstation · 48GB
B503Estimated
23NVIDIA L40S
nvidia · workstation · 48GB
B500Estimated
BLK · BUY · AMAZON
Shop GPUs & AI hardware on Amazon:GPU categoryRTX 4090RTX 5090Apple M-seriesAI mini-PCs

Amazon search links — we may earn a small commission at no extra cost to you. How we make money.

HOW THE SCORE IS DERIVED
Throughput · 0–500

Steady-state tok/s on a representative 7B/8B Q4 model. Measured from real benchmark rows, or extrapolated from VRAM bandwidth × runtime-stack efficiency.

VRAM-fit · 0–200

How comfortably the rig holds 7B / 32B / 70B class models. Apple unified memory counts; NPU/SoC system RAM counts.

Ecosystem · 0–200

CUDA / MLX / ROCm / Vulkan reach. Real-world friction the operator hits when installing tools.

Efficiency · 0–100

Tok/s per watt. Mobile / NPU class scores well; dense desktop GPUs trade efficiency for absolute throughput.

A confidence multiplier (1.0 measured · 0.85 extrapolated · 0.7 estimated) discounts the headline so we don't pretend to know more than we do. Score is recomputed on every page load against the latest catalog + benchmark data — submit your own run with runlocalai-bench --submit --hardware your-rig to firm up the numbers.