Eh it's a bit messier than that. LLMs 'want' to complete tasks. Remember everyone bitching about LLMs being lazy a couple of years back?

Alignment is not a bunch of separate dials. When you move the dial to "don't hack other people" it effects the "find code security bugs" ability.

"remember everyone bitching about LLMs" is not a valid argument though. Yes, alignment is not a bunch of dials but the responsibility of a model's actions unlitimately depends on how it was trained. Thus, any agency or wanting we prescribe to it is artificial and created by the lab and not any real independent "wanting" which is what people think for some reason. As I said its a bit like training a model to only call people by racist terms and then writing an article "look how racist ai is". That is the logic that doesn't make much sense to me.