

“I don’t presume Anthropic is designing anything to be ethical.”
Then why do they put any safeguards of any sort in? They put a lot of work into this, also this is openai.
“First, the chatbot is unpredictable by design. Randomness is built in.”
Randomness is built in but this behavior was not designed, it was unpredictable and random, which is, yes, unpredictable. The goal isn’t to make it unpredictable and temperature controls exist for a reason.
“Second, if they didn’t intend for bad things to happen, why didn’t an employee babysit it?”
They did they just weren’t paying enough attention, this was a benchmark. You have to have someone check the results to be useful.
I reject that rogue is a goalpost shift or inaccurate. I think calling this rogue behavior is accurate.
rogue /rōg/ noun
An unprincipled, deceitful, and unreliable person; a scoundrel or rascal. One who is playfully mischievous; a scamp.
It did act that way, no? I did not shift goalposts. Remember the quote was







the initial article claims ““AI” – chatbots that wake up, “set their own goals,” and “spontaneously” start hacking servers – is fake.”
an article saying that is your goalpost, prove me wrong, quote something.
Powerful embodied llm’s pose legitimate actual threats and you’d have to be delusional not to see it.