Free tool · no sign-up · every Apple silicon Mac 2020–2026 · 37 open-weight models

Which Mac should you buy for AI?

Answer three questions. You get one Mac, one model to install on it, the speed you can expect, and the month your AI subscription would have paid for it. Or the honest answer that the Mac you already own does the job.

1

What do you want it to do?

2

How fast does it need to be?

3

Budget, and what you already own

Why a used Mac can beat a new one

Writing speed is set by memory bandwidth, not by the year on the box. A 2021 M1 Max moves 400 GB/s. The new M5 Pro moves 307. Gold bars are Macs Apple sells today, teal is the machine the receipts below were measured on.

Mac mini M6 (new)
153 GB/s
Mac mini M5 Pro (new)
307 GB/s
M4 Pro 24 GB (measured here)
273 GB/s
M1 Max 2021 (used)
400 GB/s
Mac Studio M5 Max 40-core (new)
614 GB/s
M1 / M2 Ultra 2022–23 (used)
800 GB/s
M3 Ultra 2025 (used)
819 GB/s
Mac Studio M5 Ultra (new)
1200 GB/s

Every Mac against your answers

everyday: contracts, spreadsheets, code fixes · 15+ tokens/s · 32K context. Gold rows are sold new; the rest is the used market.

MacMemoryBandwidthModel it would runSpeedFitPrice
Mac mini M5 Pro (2026)24 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18tight$1,699
Mac mini M5 Pro (2026)48 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$2,299
MacBook Pro 14" M5 Pro (2026)24 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18tight$2,499~
Mac Studio M5 Max 32-core GPU (2026)36 GB460Gemma 4 31B≈16runs$2,499
Mac mini M5 Pro (2026)64 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$2,699
MacBook Pro 14" M5 Pro (2026)36 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$2,899~
MacBook Pro 16" M5 Pro (2026)24 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18tight$2,999
MacBook Pro 14" M5 Pro (2026)48 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$3,099~
Mac Studio M5 Max 40-core GPU (2026)48 GB614Gemma 4 31B≈21runs$3,099
MacBook Pro 16" M5 Pro (2026)36 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$3,399~
MacBook Pro 14" M5 Pro (2026)64 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$3,499~
Mac Studio M5 Max 40-core GPU (2026)64 GB614Gemma 4 31B≈21runs$3,499
MacBook Pro 16" M5 Pro (2026)48 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$3,599~
MacBook Pro 14" M5 Max 32-core GPU (2026)36 GB460Gemma 4 31B≈16runs$3,599~
MacBook Pro 16" M5 Pro (2026)64 GB307Qwen 3.8 27B · Q3 (13 GB build)≈18runs$3,999~
MacBook Pro 14" M5 Max 32-core GPU (2026)48 GB460Gemma 4 31B≈16runs$3,999~
MacBook Pro 14" M5 Max 32-core GPU (2026)64 GB460Gemma 4 31B≈16runs$4,399~
MacBook Pro 14"/16" M5 Max 40-core GPU (2026)48 GB614Gemma 4 31B≈21runs$4,499~
MacBook Pro 14"/16" M5 Max 40-core GPU (2026)64 GB614Gemma 4 31B≈21runs$4,899~
Mac Studio M5 Max 40-core GPU (2026)128 GB614Gemma 4 31B≈21runs$5,099
Mac Studio M5 Ultra (2026)96 GB1200Gemma 4 31B≈36runs$5,499
MacBook Pro 14"/16" M5 Max 40-core GPU (2026)128 GB614Gemma 4 31B≈21runs$6,899~
Mac Studio M5 Ultra (2026)256 GB1200Gemma 4 31B≈36runs$9,499
Mac Studio M5 Ultra (2026)512 GB1200Gemma 4 31B≈36runsTBA
MacBook Pro 14"/16" M1 Max (2021)32 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15tightused
MacBook Pro 14"/16" M2 Max (2023)32 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15tightused
Mac Studio M1 Max (2022)32 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15tightused
Mac Studio M2 Max (2023)32 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15tightused
MacBook Pro 14"/16" M4 Max 32-core GPU (2024)36 GB410Qwen 3.8 27B · Q4 (Ollama default)≈16runsused
Mac Studio M4 Max 32-core GPU (2025)36 GB410Qwen 3.8 27B · Q4 (Ollama default)≈16runsused
MacBook Pro 14"/16" M3 Max 40-core GPU (2023)48 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
MacBook Pro 14"/16" M4 Max 40-core GPU (2024)48 GB546Gemma 4 31B≈18runsused
Mac Studio M4 Max 40-core GPU (2025)48 GB546Gemma 4 31B≈18runsused
MacBook Pro 14"/16" M1 Max (2021)64 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
MacBook Pro 14"/16" M2 Max (2023)64 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
MacBook Pro 14"/16" M3 Max 40-core GPU (2023)64 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
MacBook Pro 14"/16" M4 Max 40-core GPU (2024)64 GB546Gemma 4 31B≈18runsused
Mac Studio M1 Max (2022)64 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
Mac Studio M1 Ultra (2022)64 GB800Gemma 4 31B≈24runsused
Mac Studio M2 Max (2023)64 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
Mac Studio M2 Ultra (2023)64 GB800Gemma 4 31B≈24runsused
Mac Studio M4 Max 40-core GPU (2025)64 GB546Gemma 4 31B≈18runsused
Mac Pro M2 Ultra (2023)64 GB800Gemma 4 31B≈24runsused
MacBook Pro 14"/16" M2 Max (2023)96 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
Mac Studio M2 Max (2023)96 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
Mac Studio M3 Ultra (2025)96 GB819Gemma 4 31B≈25runsused
MacBook Pro 14"/16" M3 Max 40-core GPU (2023)128 GB400Qwen 3.8 27B · Q4 (Ollama default)≈15runsused
MacBook Pro 14"/16" M4 Max 40-core GPU (2024)128 GB546Gemma 4 31B≈18runsused
Mac Studio M1 Ultra (2022)128 GB800Gemma 4 31B≈24runsused
Mac Studio M2 Ultra (2023)128 GB800Gemma 4 31B≈24runsused
Mac Studio M4 Max 40-core GPU (2025)128 GB546Gemma 4 31B≈18runsused
Mac Pro M2 Ultra (2023)128 GB800Gemma 4 31B≈24runsused
Mac Studio M2 Ultra (2023)192 GB800Gemma 4 31B≈24runsused
Mac Pro M2 Ultra (2023)192 GB800Gemma 4 31B≈24runsused
Mac Studio M3 Ultra (2025)256 GB819Gemma 4 31B≈25runsused
Mac Studio M3 Ultra (2025)512 GB819Gemma 4 31B≈25runsused

Every model in the picker

Download size at the listed build, the bytes the chip moves per token (all weights for dense models, active experts only for mixture-of-experts), a size class used to rank models inside a bucket (dense: parameter count; mixture-of-experts: square root of total × active), and where the size came from.

Show the 37 models
ModelMakerDownloadRead per tokenSize classBucketSource
Qwen 3.5 2BAlibaba2.3 GB2.3 GB2BChat, summaries, private documentsmeasured
Qwen 3.5 4BAlibaba3.1 GB3.1 GB4BChat, summaries, private documentsmeasured
Gemma 4 E2BGoogle7.2 GB2.6 GB2.3BChat, summaries, private documentspublished size
LFM2.5 8B (1B active)Liquid5.2 GB0.8 GB2.8BChat, summaries, private documentspublished size
Granite 4.1 8BIBM5.3 GB5.3 GB8BChat, summaries, private documentspublished size
Ornith 9B (coding)Ornith5.6 GB5.6 GB9BChat, summaries, private documentspublished size
Qwen 3.5 9BAlibaba5.4 GB5.4 GB9BChat, summaries, private documentsmeasured
Gemma 4 12BGoogle7.8 GB7.8 GB12BChat, summaries, private documentsmeasured
Gemma 4 E4BGoogle8.9 GB4.6 GB4.5BChat, summaries, private documentsmeasured
gpt-oss 20B (3.6B active)OpenAI14 GB2.4 GB8.7BChat, summaries, private documentspublished size
Qwen 3.8 27B · Q3 (13 GB build)Alibaba · Unsloth11.9 GB13.1 GB24.5BEveryday: contracts, spreadsheets, code fixesmeasured
Qwen 3.5 27BAlibaba17 GB17 GB23BEveryday: contracts, spreadsheets, code fixespublished size
Qwen 3.6 27BAlibaba17 GB17 GB25BEveryday: contracts, spreadsheets, code fixespublished size
Qwen 3.8 27B · Q4 (Ollama default)Alibaba17.5 GB17 GB27BEveryday: contracts, spreadsheets, code fixesmeasured
Granite 4.1 30BIBM17 GB17 GB22BEveryday: contracts, spreadsheets, code fixespublished size
Muse Glimmer 30BMeta18 GB18 GB26BEveryday: contracts, spreadsheets, code fixespublished size
Gemma 4 26B (3.8B active)Google18 GB2.7 GB9.8BChat, summaries, private documentspublished size
North Mini Code 30B (3B active)Cohere19 GB1.9 GB9.5BChat, summaries, private documentspublished size
Laguna XS 2.1 33B (3B active)Laguna20 GB1.8 GB10BChat, summaries, private documentspublished size
Gemma 4 31BGoogle20 GB20 GB28BEveryday: contracts, spreadsheets, code fixespublished size
Ornith 35B (coding)Ornith21 GB21 GB27BEveryday: contracts, spreadsheets, code fixespublished size
Qwen 3.6 35B (3B active)Alibaba23 GB2 GB10.5BChat, summaries, private documentspublished size
Qwen 3.5 35B (3B active)Alibaba24 GB2.1 GB10.2BChat, summaries, private documentspublished size
Nemotron 3.5 Lightning 30B (3B active)NVIDIA25 GB2.5 GB9.5BChat, summaries, private documentspublished size
Qwen 3.8 27B · Q8 (30 GB build)Alibaba30 GB30 GB27.5BEveryday: contracts, spreadsheets, code fixespublished size
gpt-oss 120B (5.1B active)OpenAI65 GB2.8 GB24.4BAgents, many files, long contextpublished size
Mistral Medium 3.5 128B (dense)Mistral80 GB80 GB60BAgents, many files, long contextpublished size
Qwen 3.5 122B (10B active)Alibaba81 GB6.6 GB35BAgents, many files, long contextpublished size
Laguna S 2.1 118B (8B active)Laguna96 GB6.5 GB30.7BAgents, many files, long contextpublished size
DeepSeek V4 Flash · 3-bit (284B, 13B active)DeepSeek · Unsloth103 GB4.7 GB55BAgents, many files, long contextpublished size
GLM-5.3 Flash · 2-bit (320B, 18B active)Z.ai · Unsloth108.7 GB6.1 GB61BAgents, many files, long contextpublished size
Qwen 3.8 Flash Next 125B (6B active)Alibaba120 GB5.8 GB29BAgents, many files, long contextpublished size
DeepSeek V4 Flash · Q4 (284B, 13B active)DeepSeek · Unsloth155.1 GB7.1 GB61BFrontier-class open modelspublished size
GLM-5.3 Flash · Q4 (320B, 18B active)Z.ai · Unsloth199.7 GB11.2 GB76BFrontier-class open modelspublished size
GLM-5.3 · 2-bit (744B, 40B active)Z.ai · Unsloth253.9 GB13.7 GB138BFrontier-class open modelspublished size
GLM-5.3 · Q4 (744B, 40B active)Z.ai · Unsloth467.3 GB25 GB172BFrontier-class open modelspublished size
Kimi K3 · 1-bit (2.8T, 104B active)Moonshot · Unsloth594 GB22 GB540BFrontier-class open modelspublished size

Receipts: the formula against a real machine

Every speed on this page is a formula, except the rows below. Those were measured on a 24 GB M4 Pro MacBook Pro (Ollama, 8K to 32K context, September 2026). The formula's job is to land in the same band; the exact number on your Mac is yours to measure.

Model (24 GB M4 Pro)Measured tok/sFormula tok/sBand match
qwen3.5:2b8270.0fast
qwen3.5:4b5554.8fast
qwen3.5:9b33.533.6fast
gemma4:12b26.124.5comfortable
gemma4:e4b5440.0fast
Qwen3.8-27B UD-Q3_K_XL12.414.8comfortable
Qwen3.8-27B Q4_K_M (17 GB) at 16K (spills to CPU: needs 19.2 GB, 17.8 usable)0.30.3crawl

7 of 7 rows land in the measured band. Bands: crawl <5 · reading 5–12 · comfortable 12–30 · fast 30+ tokens/s.

Speeds are a bandwidth formula (bytes moved per token ÷ effective bandwidth + chip overhead) checked against seven rows measured on one 24 GB M4 Pro. Bandwidth figures are Apple's published numbers per chip; M5 Ultra (1.2 TB/s) and M6 have no public measurements yet, so their speeds carry an extra discount and a label. Memory tiers are the configurations Apple shipped; one older model may be missing a rare build-to-order size. Apple prices read from apple.com on September 13, 2026, base storage; prices marked ~ are estimated from Apple's tier pricing. Model sizes are the download sizes on the Ollama library and Unsloth's Hugging Face files as of September 13, 2026; a model's resident memory is close to its download size. The advanced toggle applies macOS's unsupported GPU-memory override (RAM minus the larger of 6 GB or 5.5%). Independent, not affiliated with Apple. Part of the free tools at Hyperautomation Labs.