upvote
For 99.99% of people, spending 15 grand on a Mac Studio just to run Qwen 3.8 locally is a non starter.
reply
It's not the M5 Ultra itself, but the M7s or M9s that will do the damage.

99% of people will use whatever AI is free. The sophisticated, heavy users that are willing and able to pay a lot of money the ones that will be interested in controlling their inference bills.

Today, the sweet spot where an M5 Ultra makes sense is tiny. But we might expect that to grow a lot.

reply
Anthropic is reporting 100 Billion ARR.

Even if you could get a frontier model, you would not be able to run it on any Mac. So speculating on what M7 or M9 will achieve in 5 years (if we even still exist) seems pointless.

reply
Do you need a frontier model to write emails, check your calendar, search the web?

I don’t think Apple is going to lie down and cede AI to the cloud.

reply
> Do you need a frontier model to write emails, check your calendar, search the web?

How about writing mail to President and senators on AI doomsday scenario if frontier labs do not pace themselves?

That mini model on mac mini would scared to hell to do such thing. It need that rugged frontier model to speak truth to power.

reply
True, but do you need a 10,000 dollar machine to do so?

I would love to find an excuse to buy a 10,000 dollar machine! But I cant find one yet. My current cloud bill is in excess of 400 USD per month. Just can't achieve frontier model capabilities locally.

reply
I think in 5 years (aka before OpenAI can pay off all its debt) a lot of AI will be running on your iPhone
reply
People thought that 5 years ago, too. OpenAI has competitors to worry about, but Apple isn't one of them.
reply
Not for some of those well-connected high school, college kids, whose parents are rolling in it, the next generation of talented Steve Jobs, Bill Gates, Zuckerberg’s and others are coming up… Guess what they’re gonna get for Christmas?
reply
Yes, but in 7 years?
reply
In 7 years we will probably all be living under ground fighting skynet with plasma rifles made out of old microwave parts
reply
I’ve watched Terminator 2 extensively. I am prepared
reply
You will. Some of us are going to be already ground into dust that the microwave parts are made out of. Others will be locked into our communism cubes with our daily allotment of entertainment and sustinece. let out into the sunlight for only 30 minutes per day.
reply
at 15 tok/s
reply
And nightmare fuel is just what they'll be selling at the UN this week, for this very reason.

Sam's address will probably be more riveting, imaginative, and terrifying than the last couple of Terminator screenplays. Legislators will lobby him to write the laws for them, and the ghost of Harlan Ellison will threaten to sue him.

reply
I don't see how that math works? This is a $15k rig under benchmark and per the results it competes very acceptably against... one consumer GPU.

I really don't see who buys this, except people who want the Studio for some other reason. But nothing in the story says you want to fill racks with these instead of Blackwell or TPU parts; it's not even close.

reply
Your math is correct, but it’s math based on today’s economics.

Think of a company like Apple moving onto your turf. They’re not going to cede AI to the cloud. They want their part of the pie.

So in 7 years, how much AI will be handled locally on your iPhone. And will you have repaid all the debt on your balance sheet before Apple eats your lunch

reply
I think most important thing is that Nvidia doesnt want to give 100% of market to frontier AI labs either.

It's way too easy for 1T+ frontier labs to ditch Nvidia. So Nvidia will also put effort to make sure there are open weights models and local hardware available.

And Apple will benefit from this too.

reply
There is zero chance that an LLM approximating a modern frontier model is going to be running on a phone in the next decade. Even if you grant that you could stack enough DRAM dies on top of each other in the package, that would be a three order of magnitude improvement in power efficiency just for the compute.
reply