upvote
not as crazy a config, but I got a 128GB M5max the day before the price increase for $2K less. still proud of the call to buy it.
reply
Also proud of having spent $2.5k on a used 128G M1 Ultra back in September 2024.

Despite being outdated in terms of compute, it stills let me run very good recent models locally, with Deepseek V4 Flash 0731 being the greatest one right now, and hopefully Qwen 3.8 Flash will also fit well when it is released tomorrow!

The M1 ultra definitely leaves to be desired in terms of its token speeds, but I think 20 tps generation and ~200 tps prompt processing (which is what I get with DSv4 flash), is already enough to do a lot of serious work when you combine with the decent prompt caching provided by llama.cpp.

reply
You're killing me here lol
reply