The question mark in my mind over the technological superiority is whether the additional volume of data they see due to capturing the top of the market allows them to do recursive self-improvement in a way nobody else can match, before any of the other labs can figure it out. That's the only runaway outcome I can see.
If you have exponentially increasing use of your harness, then it's true that every day you capture exponentially more data, but it's also true that every day exponentially more data will slip through the cracks of your would-be monopoly and that data arrives at your competitors via various channels (competitor harnesses, subsidized reselling, etc)
The very exponential that you are relying on to give you runaway improvement is also giving exponentially increasing data to your competitors. All else being equal your competitors stay a step behind but you never develop a monopoly either. That's the best case for Anthropic/OpenAI. In reality, training data is just one variable, exponentials don't last forever, and your competitors will get better at capturing a bigger slice of training data.
If user data would become such a key ingredient (which it might, i actually remember noam shazeer talking about the importance of user data), i think chinese labs can still get it from china, as keep in mind it ahs a billion people behind the great firewall banned from using us llms. And btw broadly for any gap like this, you really gotta consider that if its becoming a bottleneck, chinese labs will find a way to buy it from one of the labs unless theres strict regulation at the government level
But is that data good? That's the question. As in, is my usage at work:
a) indicative of problems that aren't already out there in the wild? (no) b) are the responses I'm getting so good and novel that the model can improve itself? (no)
It's the garbage in garbage out idea, just scaled up. If the model gave a bad answer, and I didn't catch it, and you now train on that I/O pair (my perhaps crappy prompt, the bad output), then you're not going to improve anything.
It seems like the user response rating mechanism might be a valuable signal
Yes, RSI seems to be the new AI industry McGuffin of 2026, just as agentic capability has become table stakes and scaremongering has become a punchline.