upvote
By this kind of logic we'd paying something like ~$10,000/month for our internet connections. The perceived value needs to exceed the cost but that does not make it the only factor to consider price with.

What makes this expensive & sell well is it's not very fungible at the moment. Where else are you going to get 512 GB of high speed memory with a well supported accelerator attached that you can throw in the corner of anyone's home and not really have them notice? There are plenty of lesser options, plenty of noiser/power hungry options, plenty of harder to support options, but not really something in direct competition at the moment. Even the next rounds of the integrated AMD/Nvidia solutions are only targeting 196 GB of much slower memory and compute.

reply
I think the difference besides the supply crunch there is that everyone connected gets the ~same internet, just faster or slower. Quantitative, not qualitative difference. On the other hand, a computer that can run Gemma 4 8B versus one that can run DeepSeek Flash are different enough experiences that I'd say they're effectively a difference in kind. It's been a bit since we had such serious stratification in outright capability in computing, rather than just how long it takes to get something done, or how many of something it can serve at once. In the early 90s, I think there were a lot more of those "this computer can do this thing, this one just can't" scenarios.

Closest competition I see right now are stacks of 2-4 connected DGX Sparks, similar lowish speed high mem, and about the same cost/gig.

reply
If it were about capability instead of speed then you're welcome to pay me $5,000 for DeepSeek Flash running on an SSD :D. For $10,000 I'll even give you something which can run 700 GB models comfortably from RAM - not that the speed should be worth much.
reply
Haha right, well, usable speed. And I guess there are some parallels with internet, the internet is technically usable with HughesNet, but people used to fiber would probably consider it unusable.
reply
I like this train of thought. The inverse is saying that the cost of this computer is the value we give away to AI companies by doing compute on their servers with our data. And to take it another way, is the value to you, the cost of a small used car?
reply
> And to take it another way, is the value to you, the cost of a small used car?

For me personally, not quite that valuable yet, but I think it's getting there quickly. Deepseek V4 Flash massively increased the value of local AI to me, to the point where it's displaced most of my Claude Code usage, its upcoming vision enabled version should bump it further, and it's only going to get better from there.

It's a lot faster, but a lot of it is also feeling free to discuss things I wouldn't be comfortable sending to Claude, with the idea that that info is now theirs in perpetuity. I got my genome fully sequenced recently (it's cheap now!), and I get a battery of blood tests every year. Wouldn't do processing on any of that with Claude, but local AI? Totally great.

And if I was running a company with a large cloud AI bill, I'd probably buy a wheelbarrow full of these macs. Cheaper, but also a more solid/predictable base to build on.

reply
Nvidia is definitely working on a AI Rack machine fully built for on-prem uses for companies.
reply
Could you share what you did to find events to do? That sounds really cool
reply
Sure, it’s pretty dumb/naive implementation, would need to be a lot more efficient to scale. Basically had Claude write polite/low touch dumb crawlers for a bunch of local sites (library, local events spaces, luma, theaters, maker spaces) and whip up a little frontend to let our family and friends manage a little text description of what their family members like, constraints, that sort of thing. Once a week, the crawlers look for new events, add to a db, and then run through the list and ask our local LLM to grade each event, given the text description and constraints. Take the top 20ish for the following two weekends and email out. It’s been super helpful, lowers the activation energy to go to more local events.
reply
Do you really need a powerful Mac to do this? Wouldnt this work on a base model or even an old Windows computer?
reply
Probably don’t really need it, just makes it easier. I tried it with some potato class models first and they would ignore or mishandle constraints, especially when vaguely worded. Sometimes it would try to send us to events that were obviously in the middle of the school day, and I’d look at their reasoning traces and there’d be some really boneheaded mistakes in there. If it reliably sends you a decent fraction of garbage reccs, the emails stop getting opened, my SO has no patience for that.
reply
[flagged]
reply
Hey, a high school jock from 1982 wandered in to the computer lab...
reply
[flagged]
reply
Weird swipe, I'm not trying to sell it to you, this is just where I see it going, and why these things are going to sell (and why 512 gig M3 Ultra Mac Studios have skyrocketed in demand/price). It's not been an empty promise, in that it's been providing lots of very concrete value to us already.
reply
Members of my C-suite are doing this today on a raspberry pi, so the promise is not empty. Key difference is they aren't using local AI to do it.

For the M5 Ultra, I suspect it would be valuable for someone who wants to achieve all the above and more, but with local AI due to data privacy concerns, and also not regulated data that comes with lots of other requirements solved by more traditional approaches.

Three possibilities:

1. The type of person who deals with lots of intellectual property using expensive Mac-only desktop applications that aren't meant for servers, whose mind has formed positive associations with the term "Apple Intelligence", whose values overlap with Apple's lawyer's values, who actually stands to profit from having a Mac that's more powerful than anyone else's Mac, whose long-term goals are not impacted by planned obsolecense on a piece of computer hardware costing over $25k ($50k after 1TB SSD add-on).

2. Trust fund beneficiary who wants to show off, LARP as #1, prime target for Apple's marketing.

3. 2026 kit for billionare-class iPad babies. All brain rot content is 100% local AI-generated. Never have to speak to your children again. A true "we have dead internet theory at home" machine.

reply
If you actually want to learn, highly recommend a visit to https://old.reddit.com/r/LocalLLaMA/. People there are running stacks of 2-4x DGX Sparks to get to similar levels of memory, at similar cost.
reply
What exactly is empty about any of what he wrote?
reply