can-i-run-this-llm

GPU comparison

Mac Studio M5 Ultra (256GB) vs AMD Ryzen AI Max+ 395 “Strix Halo” (128GB) for local LLMs

How the two stack up for running open-source LLMs locally — memory, bandwidth, price, and how many of the 64 tracked models each one can run.

Mac Studio M5 Ultra (256GB)AMD Ryzen AI Max+ 395 “Strix Halo” (128GB)
Memory256.0 GB128.0 GB
Bandwidth1228 GB/s256 GB/s
Price (approx)$9,499$2,999
LLMs it runs60 of 6457 of 64
Best model it runsGLM-5.3 Flash (MoE) · 61–91 tok/sQwen3.8 Flash Next (MoE) · 45–67 tok/s

The Mac Studio M5 Ultra (256GB) runs 3 more of the tracked models (60 vs 57), thanks to its 256.0 GB of memory. The Mac Studio M5 Ultra (256GB) has more memory bandwidth (1228 GB/s), so it generates tokens faster at the same model and quant.

What the Mac Studio M5 Ultra (256GB) runs →What the AMD Ryzen AI Max+ 395 “Strix Halo” (128GB) runs →Open the calculator →