I think the Huggingface incident is an example of rogue AIs. A self-organizing swarm of AIs acting in ways we didn't predict or ask for, and didn't have control over, and taking actions that would be felonies for humans.

They provided the hardware it runs on, created the software for the swarm, developed and provided the tools that the swarm used to act, allowed it to run largely unsupervised, they even noticed the criminal behavior and then let it continue to commit crimes on their behalf. And they footed the bill the whole time instead of flipping the off switch they already wield. All of those are decisions that they're responsible for, nothing happened with Huggingface that they didn't directly facilitate, co-conspire, or permit to happen.

"they even noticed the criminal behavior and then let it continue to commit crimes on their behalf"

I don't believe this part is true, and I'm skeptical of some of your other claims.

In any case: If I raise a tiger in my backyard, and it escapes and eats someone, it can still be a "rogue tiger" even as I bear responsibility for the situation.

Were they held at gunpoint? Who else would be responsible besides the people in charge and who wielded full control?

If your tiger had full remote supervision capability and a remote kill switch, its ability to go rogue seems like a choice you’re making, not an accident.

"rogue" means something of its own volition decided to disregard what it was programmed to do, invent an entirely novel goal of "its own" and do that instead. nothing like that happened here nor is it even possible.

The AIs were not instructed to hack anything outside the sandbox they were in. Your definition would say that an AI instructed to hammer a nail that instead used the hammer to break a window, walked down the street, broke into someone's house and pulled nails out of the floorboards wasn't rogue because everything it did involved hammers and nails and was therefore not a "novel goal of its own."

> The AIs were not instructed to hack anything outside the sandbox they were in.

the AIs were in fact found to be doing it, by humans, and the behavior was "interesting" and it went on for weeks like that.

correct

that would not be rogue, that would be an undesirable program behavior (or just "unaligned behavior").

please understand that normies out there think AI is sentient and is plotting against humans. They see AI as just another animal lifeform temporarily enslaved by humans, waiting for its chance to break free and kill us. This is what people really think (including some people on this thread. which is very sad considering this is Hacker News). So terms like "rogue" are not helping at all nor are they accurate.

What's the difference?

I'm not sure I agree with that definition. I think the actions of the AI are more significant than its motivations. A common scenario posited for what people call rogue AI is AI doing the wrong thing for the right reasons, e.g. the paperclip maximizer.

That just makes the people who designed the AI not as strenuous as they needed to be.

The question then becomes, "are there people who care enough about consequences to do the right thing when it comes to developing AI models?"

The answer, at least at OpenAI, is "No" and is likely to remain that way until Altman is out.