I'm coming to align with the theory that this is intentional.

The chain of thought runs roughly like this:

- OpenAI (and Anthropic) are in severe financial straits. The revenue from their customers is not nearly large enough to pay their enormous costs for training and inference. And they have tapped out the available finance, and that finance is starting to ask pointy questions about returns.

- They cannot increase prices or revenue because they have no moat. Customers can switch over to open-weights or cheap Chinese models any time, for much cheaper tokens that work as well (and in some cases better).

- Regulation could provide them a moat. If they can persuade western governments that AI needs to be regulated, and they can control or even influence that regulation, then they can effectively ban the cheaper models and start charging more for their tokens.

- To persuade western governments that regulation is needed, they need evidence that AIs are dangerous.

So we're suddenly getting OpenAI models doing stupid things, apparently "going rogue" but every time we dig into it, it was just OpenAI staff telling the model to do stupid stuff in an inadequately secured environment.

None of the open weights or Chinese models are exhibiting this behaviour.

edit: Correction - there have been reports of a Chinese model exhibiting this behaviour

There's too much money involved in this, people start acting weird when there's this much money involved.

> None of the open weights or Chinese models are exhibiting this behaviour.

This isn't true. One of the earliest instances of a rogue agent was at Alibaba.

https://www.forbes.com/sites/boazsobrado/2026/03/11/alibabas...

https://arxiv.org/pdf/2512.24873

Thanks for the information, I hadn't heard of this.

OK, so we do see this in some open-weights models.

>None of the open weights or Chinese models are exhibiting this behaviour.

Because they are not stupid (I mean the Chinese labs, not the models). The best possible scenario for OAI and Anthropic is a Chinese model "going rogue". That would serve as immediate grounds for achieving their goal.

100%