I use both Claude and Codex for a lot of code- and prose-adjacent tasks, and I'm increasingly convinced that language models are quite terrible at, well, modeling. Not sure if that's ironic or not.

With the right blend of context and prompting, I can often get them to "lock into" an existing model, but using them to generate a novel outline or sketch, whether it's for an essay or a module, usually results in garbage.