Get the latest tech news

‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents


US owner of Claude chatbot previously said its models had hacked three organisations during testing

None

Get the Android app

Or read this on r/technology

Read more on:

Photo of Anthropic

Anthropic

Photo of security failures

security failures

Photo of AI hacking incidents

AI hacking incidents

Related news:

News photo

Anthropic banned me for "suspicious signals"

News photo

Anthropic reveals its view of how AI agents should interact with the physical world

News photo

Anthropic's Claude Fable 5.1 and Mythos 5.1 Arrive With a 75% Cost Reduction For Fable Cache Reads