Reasoning tokens are a way to escape autoregressive woes. The model can generate a draft, then ponder on it, and use this to generate a final version

They’re a way to mitigate it. It still writes like an LLM and everyone can see it.