The thinking traces on some Chinese models just output the full response in the thinking trace, then output it again to the user, which is redundant.
The thinking traces on some Chinese models just output the full response in the thinking trace, then output it again to the user, which is redundant.