The main reason we don't see much quality degradation in LLM writing output is because they're already poor writers. This is the load bearing reason.
I was bulding a small interpreter and writing an article in ~markdown yesterday with Fable. And while it codes like a pro, it writes like a sixth grader.
Let's see how these watermarking stats hold up if/when llms start writing well.