What is the moat? The time it takes for AI to rewrite an efficient inference stack for a new model? Considering most LLMs follow a similar architecture, adapting to a new model shouldn't take that much time.
There is no moat. At the moment, all of these companies are burning money to gain mindshare and market share. That's what Thinking Machines is doing; they're not looking for a business model.
I don't know why people keep saying there's no moat. There's no moat. Having a FUCK ton of money to train these gigantic fucking models and retain the brains to make it happen is a moat.
You're not going to train one using a VPS from LowEndBox.