← r/LocalLLaMA
▲
195
+2
41👁
r/LocalLLaMA · u/dreamingwell · 14d ago

M5 Ultra 80Core GLM-5.3-Flash on DwarfStar Speeds

post image

I've been playing around with various models on the M5 Ultra 256GB 80-core Mac Studio. These are the results over many rounds of agentic inferencing.

I'm happy with the performance. Glad to have the large amount of RAM. But it does feel like the GPU is underpowered for this amount of RAM. I'm wondering if a 512GB unit for AI inference makes sense at all - because the GPU will be the clear bottleneck.

197 0 195 10/3 04:42 10/9 05:04 UTC
scorecomments41 sightings
first seen 2026-10-03 04:42 UTClast seen 2026-10-09 05:04 UTCscore then 193score now 195gained +2sightings 41
open on reddit ↗ 💬 114 (+4)