I've been holding out, because I think my next purchase will be a Studio with an Ultra Chip in it. I'm wanting it to be a "forever" server, so I'm holding out while I can.
If running LLMs locally matters, it’s hard to imagine a “forever machine” existing in anything less than 5-10 years, probably more. This stuff is just evolving so rapidly. Buying a “forever machine” today might be like buying a “forever GPU” in 2003.
Supply rumours are next year we see an M7 AI-focused chip with large inference performance upgrades. It's unlikely we'll see heavy upgrades in other areas. If you care about AI, it's worth waiting. If you don't, pull the trigger now. RAM constraints are likely to get worse next year. Or wait 2-3 years and prices should be back to Earth (plus newer and even better chips).
Yeah, but how many years until 128GB+ is attainable by mere mortals again? My 2021 home server build was 64GB of RAM. My 2025 build was 32GB and zram. :/
I just wait for a cycle or two where the leaps and bounds are more like hops and steps. So if the M7 Ultra improves inference by 2x over the M5, but the M9 Ultra only improves by 1.2x over the M7, that's my signal to buy. Unfortunately they haven't slowed down yet.