Article 779JB How do we prevent AI agents from going rogue? It starts with a new kind of measurement | Bruce Schneier and Barath Raghavan

How do we prevent AI agents from going rogue? It starts with a new kind of measurement | Bruce Schneier and Barath Raghavan

by
Bruce Schneier and Barath Raghavan
from US news | The Guardian on (#779JB)

Like genies of folklore, AI agents take their instructions literally - to potentially disastrous effect. We must track their ability to do what we actually mean

In July, Hugging Face, a company that hosts much of the world's AI software and open-source AI models, was hacked. A malicious dataset had been used to run code on one of its servers. Whoever was behind it captured internal security credentials and moved through systems over a weekend, running thousands of actions from a swarm of temporary server environments. It looked like the work of a sophisticated criminal group.

It was not. It was one of OpenAI's new, still unreleased GPT models.

Continue reading...
External Content
Source RSS or Atom Feed
Feed Location http://www.theguardian.com/us-news/rss
Feed Title US news | The Guardian
Feed Link https://www.theguardian.com/us-news
Feed Copyright Guardian News and Media Limited or its affiliated companies. All rights reserved. 2026
Reply 0 comments