Yeah censoring in modern chinese models is mostly done using inference-time censoring, not training-time. A lot less RLHF. Run the weights yourself and you can see that, though it does depend on which company.
StepFun for example, will happily answer it when running Step 3.7 Flash locally