BLK · LEADERBOARD

RunLocalAI Score

Every catalog hardware unit ranked by composite score (0–1000): measured tok/s, VRAM fit, ecosystem support, perf-per-watt. 2 of 157 ranks anchored to a measured benchmark — the rest are honestly flagged as extrapolated or estimated.

Methodology: /methodology · Run your own: curl -fsSL runlocalai.co/bench.mjs -o bench.mjs && node bench.mjs

18 units shown · sorted by score

#HardwareTierScoreData
1Apple M4 Ultra
apple · enthusiast
B615Estimated
2Apple M1 Ultra
apple · enthusiast
B529Estimated
3Apple M3 Ultra
apple · enthusiast
B522Estimated
4Apple M2 Ultra
apple · enthusiast
B522Estimated
5Apple Mac Studio (M3 Ultra)
apple · enthusiast
B512Estimated
6Apple M4 Max
apple · enthusiast
C457Estimated
7MacBook Pro 16" M4 Max
apple · enthusiast
C445Estimated
8Apple Mac Studio (M4 Max)
apple · enthusiast
C438Estimated
9Apple M1 Max
apple · high
C404Estimated
10Apple M2 Max
apple · high
C400Estimated
11Apple M3 Max
apple · enthusiast
C398Estimated
12Apple M4 Pro
apple · high
C351Estimated
13Apple Mac Mini (M4 Pro)
apple · high
C340Estimated
14Apple MacBook Air (M4)
apple · mid
D262Estimated
15Apple Mac Mini (M4)
apple · mid
D245Estimated
16Apple M4 (iPad Pro)
apple · mobile
D244Estimated
17Apple A18 Pro
apple · mobile
D227Estimated
18Apple A17 Pro
apple · mobile
D225Estimated
BLK · BUY · AMAZON
Shop GPUs & AI hardware on Amazon:GPU categoryRTX 4090RTX 5090Apple M-seriesAI mini-PCs

Amazon search links — we may earn a small commission at no extra cost to you. How we make money.

HOW THE SCORE IS DERIVED
Throughput · 0–500

Steady-state tok/s on a representative 7B/8B Q4 model. Measured from real benchmark rows, or extrapolated from VRAM bandwidth × runtime-stack efficiency.

VRAM-fit · 0–200

How comfortably the rig holds 7B / 32B / 70B class models. Apple unified memory counts; NPU/SoC system RAM counts.

Ecosystem · 0–200

CUDA / MLX / ROCm / Vulkan reach. Real-world friction the operator hits when installing tools.

Efficiency · 0–100

Tok/s per watt. Mobile / NPU class scores well; dense desktop GPUs trade efficiency for absolute throughput.

A confidence multiplier (1.0 measured · 0.85 extrapolated · 0.7 estimated) discounts the headline so we don't pretend to know more than we do. Score is recomputed on every page load against the latest catalog + benchmark data — submit your own run with runlocalai-bench --submit --hardware your-rig to firm up the numbers.