Smaller customers are also less able to develop their own hardware and threaten NVidia's business.
I wonder if Nvidia is also trying to cover themselves against a market crash the bankrupts the AI labs but leaves the tech standing? (Similar to the dotcom crash or the big railroad crash back in the day.) If Anthropic and OpenAI struggle, Nvidia can always sell local inference hardware. But local inference hardware often has a lower utilization, so you need more GPUs for the same number of tokens.
They can't ride the hyperscaler gravy train forever; at some point between Google, AMD, and Apple NVIDIA is going to lose its monopoly on serving large customers.
At that point, it would be useful if a few open models existed which were only a couple of months behind the frontier.
But it's very important that the open models never be TOO good, because the AI companies are buying compute on the assumption that their software will add value. If it becomes a commodity business with frontier open models, NVIDIA won't be able to get away with such a crazy markup.
Of course they would. You need hardware to run the model. nVidia sets the floor.
I think NVIDIA does want small open models running on-prem to explode as a market! Lots of smaller GPU installations for companies who wisely want on-prem inference.
Of course NVIDIA will also keep making a ton of money selling to hyper scalers, but not forever: Chinese chips are getting better, Google, Microsoft, Amazon, etc. designing their own inference chips.
NVIDIA is handling this brilliantly.
No one wants to pay the Nvidia tax
But, not really at the current technologies. Kimi and GLM are fucking awesome, but I don’t have 3TB of VRAM to run them, and I don’t expect to even when ram prices drop.
So now you’re back to the scaling issue before talking about power and compute distribution.
I can see one of Nvidia's biggest fears is the inference hardware becoming commoditised.
What Nvidia has the market cornered on is flexible GPU architectures. Nobody else has stepped up to the plate on that, and it's how Nvidia will butter their bread with robotics and future model training efforts.
are they actually suppressing the western open models?
china doesn't give a fuck either way
imo, they see the weakness emerging at the intersection of all the labs, everybody knew there was no moat, so they're gonna control its direction and basically tell the Jev guys what they want them to work on
They are just trying to grow the pie because they have nobody else competing for slices.
My guess is they want as many frontier models using their chips as possible. The only threat to their business is companies making their own chips which Google does and the others are working toward. The last thing they want is only 3 frontier labs who are all not buying NVIDIA.
Hugging Face - distribution for model that you can run on your local Nvidia Spark
Neoclouds - Nvidia setup a 500 billion investment fund with Wall Street so Nvidia can sell chip and this news about buying a LLM start up. it seem like another customer for Nvidia.
correct me if i'm wrong but i remember i saw an interview with Jensen where he want more company to have their own model and country to have their own LLM model.
Nvidia doesn't make money from the gold rush. Nvidia made money from selling the shovels.
But we’re getting ahead of ourselves
similar to how eventually AWS started making their own chips for data centers, and Apple did that for their hardware, it's not a ridiculous thing to plan for the AI companies to start making their own chips to optimize for their use cases and cut out the middleman for margins.