BLK · LEADERBOARD

RunLocalAI Score

Every catalog hardware unit ranked by composite score (0–1000): measured tok/s, VRAM fit, ecosystem support, perf-per-watt. 2 of 157 ranks anchored to a measured benchmark — the rest are honestly flagged as extrapolated or estimated.

Methodology: /methodology · Run your own: curl -fsSL runlocalai.co/bench.mjs -o bench.mjs && node bench.mjs

54 units shown · sorted by score

#HardwareTierScoreData
1NVIDIA GeForce RTX 3080 16GB (Mobile)
nvidia · high · 16GB · 27 bench
C487Measured
2NVIDIA RTX A6000 (Ampere)
nvidia · workstation · 48GB
C477Estimated
3NVIDIA GeForce RTX 5070 Ti
nvidia · high · 16GB
C477Estimated
4NVIDIA RTX A5000
nvidia · workstation · 24GB
C468Estimated
5NVIDIA A40
nvidia · workstation · 48GB
C458Estimated
6Apple M4 Max
apple · enthusiast
C457Estimated
7NVIDIA GeForce RTX 3080 Ti
nvidia · enthusiast · 12GB
C456Estimated
8NVIDIA GeForce RTX 3080 12GB
nvidia · high · 12GB
C456Estimated
9NVIDIA RTX PRO 4000 Blackwell
nvidia · workstation · 24GB
C455Estimated
10MacBook Pro 16" M4 Max
apple · enthusiast
C445Estimated
11Apple Mac Studio (M4 Max)
apple · enthusiast
C438Estimated
12NVIDIA GeForce RTX 4080 Super
nvidia · high · 16GB
C433Estimated
13NVIDIA GeForce RTX 4080
nvidia · high · 16GB
C428Estimated
14AMD Radeon RX 7900 XTX
amd · enthusiast · 24GB
C420Estimated
15NVIDIA GeForce RTX 4070 Ti Super
nvidia · high · 16GB
C418Estimated
16NVIDIA RTX 5000 Ada Generation
nvidia · workstation · 32GB
C414Estimated
17NVIDIA RTX 2080 Ti 22GB (China-mod)
nvidia · mid · 22GB
C405Estimated
18Apple M1 Max
apple · high
C404Estimated
19Apple M2 Max
apple · high
C400Estimated
20NVIDIA GeForce RTX 4090 Mobile
nvidia · enthusiast · 16GB
C400Estimated
21NVIDIA GeForce RTX 5070
nvidia · mid · 12GB
C399Estimated
22Apple M3 Max
apple · enthusiast
C398Estimated
23NVIDIA GeForce RTX 3080 10GB
nvidia · high · 10GB
C397Estimated
24AMD Radeon RX 7900 XT
amd · enthusiast · 20GB
C365Estimated
25NVIDIA GeForce RTX 5060 Ti 16GB
nvidia · mid · 16GB
C364Estimated
26NVIDIA GeForce RTX 2080 Ti
nvidia · enthusiast · 11GB
C363Estimated
27NVIDIA L4
nvidia · workstation · 24GB
C360Estimated
28NVIDIA GeForce RTX 3070 Ti
nvidia · high · 8GB
C358Estimated
29NVIDIA GeForce RTX 4070
nvidia · mid · 12GB
C356Estimated
30NVIDIA GeForce RTX 4070 Super
nvidia · mid · 12GB
C355Estimated
31Apple M4 Pro
apple · high
C351Estimated
32NVIDIA GeForce RTX 4070 Ti
nvidia · high · 12GB
C351Estimated
33Apple Mac Mini (M4 Pro)
apple · high
C340Estimated
34AMD Radeon RX 9060 XT
amd · mid · 16GB
C339Estimated
35NVIDIA GeForce RTX 5070 Laptop GPU
nvidia · high · 12GB
C337Estimated
36AMD Radeon RX 9070 XT
amd · high · 16GB
C332Estimated
37AMD Radeon RX 9070
amd · high · 16GB
C332Estimated
38NVIDIA GeForce RTX 2080 Super
nvidia · high · 8GB
C330Estimated
39AMD Radeon RX 7800 XT
amd · high · 16GB
C329Estimated
40NVIDIA GeForce GTX 1080 Ti
nvidia · high · 11GB
C327Estimated
41NVIDIA GeForce RTX 5060
nvidia · entry · 8GB
C326Estimated
42NVIDIA GeForce RTX 2060 Super
nvidia · mid · 8GB
C323Estimated
43NVIDIA GeForce RTX 2070
nvidia · high · 8GB
C323Estimated
44NVIDIA GeForce RTX 5060 Ti 8GB
nvidia · mid · 8GB
C322Estimated
45NVIDIA GeForce RTX 3060 Ti
nvidia · high · 8GB
C321Estimated
46NVIDIA GeForce RTX 4060 Ti 16GB
nvidia · mid · 16GB
C320Estimated
47NVIDIA GeForce RTX 3060 12GB
nvidia · mid · 12GB
C319Estimated
48NVIDIA GeForce RTX 3070
nvidia · mid · 8GB
C319Estimated
49AMD Radeon RX 7900 GRE
amd · high · 16GB
C319Estimated
50NVIDIA GeForce RTX 2070 Super
nvidia · high · 8GB
C319Estimated
51AMD Radeon RX 6950 XT
amd · enthusiast · 16GB
C316Estimated
52AMD Radeon RX 6800
amd · high · 16GB
C304Estimated
53AMD Radeon RX 6800 XT
amd · enthusiast · 16GB
C302Estimated
54AMD Radeon RX 6900 XT
amd · enthusiast · 16GB
C302Estimated
BLK · BUY · AMAZON
Shop GPUs & AI hardware on Amazon:GPU categoryRTX 4090RTX 5090Apple M-seriesAI mini-PCs

Amazon search links — we may earn a small commission at no extra cost to you. How we make money.

HOW THE SCORE IS DERIVED
Throughput · 0–500

Steady-state tok/s on a representative 7B/8B Q4 model. Measured from real benchmark rows, or extrapolated from VRAM bandwidth × runtime-stack efficiency.

VRAM-fit · 0–200

How comfortably the rig holds 7B / 32B / 70B class models. Apple unified memory counts; NPU/SoC system RAM counts.

Ecosystem · 0–200

CUDA / MLX / ROCm / Vulkan reach. Real-world friction the operator hits when installing tools.

Efficiency · 0–100

Tok/s per watt. Mobile / NPU class scores well; dense desktop GPUs trade efficiency for absolute throughput.

A confidence multiplier (1.0 measured · 0.85 extrapolated · 0.7 estimated) discounts the headline so we don't pretend to know more than we do. Score is recomputed on every page load against the latest catalog + benchmark data — submit your own run with runlocalai-bench --submit --hardware your-rig to firm up the numbers.