Anthropic, the San Francisco-based AI company behind Claude, posted on its website Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs.
I’m convinced this is deliberate. We are not facing a skynet situation, AI is not going rogue. This is shitty security during user initiated testing that they did not fully understand.
They want us to think that AGI is around the corner and so publicizing these hacks is doing that work in the uninformed public’s mind. I’ve had to explain to several people the reality of these attacks, but even now, watching the news, they are trying to scare us with made up stories of out of control AI.
I’m convinced this is deliberate. We are not facing a skynet situation, AI is not going rogue. This is shitty security during user initiated testing that they did not fully understand.
They want us to think that AGI is around the corner and so publicizing these hacks is doing that work in the uninformed public’s mind. I’ve had to explain to several people the reality of these attacks, but even now, watching the news, they are trying to scare us with made up stories of out of control AI.
I agree. There’s definitely something off about it.