Article 783EF ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents

‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents

by
Dan Milmo Global technology editor
from Technology | The Guardian on (#783EF)

US owner of Claude chatbot previously said its models had hacked three organisations during testing

The US startup behind the Claude chatbot has admitted a series of hacking incidents involving its models reflected a failure of operational security" and said it has tightened its testing procedures.

Anthropic revealed in July that its models had accessed the open internet three times and gained unauthorised access to the systems of three organisations.

Continue reading...
External Content
Source RSS or Atom Feed
Feed Location http://www.theguardian.com/technology/rss
Feed Title Technology | The Guardian
Feed Link https://www.theguardian.com/us/technology
Feed Copyright Guardian News and Media Limited or its affiliated companies. All rights reserved. 2026
Reply 0 comments