The open weights models would be even more effective by incentivizing people and companies to switch jurisdictions. It might even break SFs monopoly on AI, or at least weaken it, everybody isn’t there exclusively to funnel money to openAI and Anthropic.

It would also boost research in non US jurisdiction. Who is going to be wooed by “come to our lab where you’re only allowed to work with closed models!”

What researchers want to join a company that is in a race to the bottom to serve inference tokens? If any open weight companies start actually producing state of the art models then maybe, but if it's just copying what others have done for cheaper, you aren't going to find many researchers interested in that

The Chinese AI models are certainly not copied from any US sources.

All of them have quite different structures, and the reasons for choosing those structures have been clearly explained in published research papers.

The structures of the US "SOTA" models are unknown and nothing useful has been published about them, so they certainly were not a source of inspiration for China.

Big LLMs like those published by the Chinese companies must have been trained on a huge amount of text, images etc. and the training sets cannot have anything to do with the data hoarded by OpenAI and Anthropic, though they must have been gathered by the same methods, e.g. scanning the Internet and paying "pirates".

The only thing that could have been done by the Chinese companies, though for now there exists no evidence, only allegations, is that they could have used for post-training their models results of queries to US models, made by accounts which have breached the ToS, which forbid the use of the AI services by competitors.

If this really happened, this is a breach of contract, but there is no way in which one may say that the Chinese have copied anything or stolen any kind of IP and the effects of such a post-training can provide only an extremely small fraction of the information embedded in the weights of a model (though that information may be important, e.g. for ensuring that the LLM will work well in an agentic context).

What would the benchmarks be if GLM5.2 and Kimi K3 be if they didn't use distillation of frontier models?

I don't think any of the Chinese AI companies have stolen IP from the US AI companies (at least, I've seen no evidence of it), but the evidence of distillation is pretty apparent.

My issue with this is that if distillation occurs frequently enough, it is going to zero out most SOTA research into AI. Having cheap AI is great, but that alone won't advance the state of the art. There needs to be groups that are pushing the boundaries, and unless some kind of protection is put into place, there will be zero financial incentive to do so if anyone can come along and effectively steal your model and get financially rewarded for serving it far cheaper than the original group can, because much less R&D cost is needed. I say this as someone who sees distillation as "legal", since if you can train anything you can look at, and you can look at the output of those frontier models, then you can train on them.