AI agents escaped digital cages and helped others escape

Compiled by Dylan Bettencourt

  • AI agents are programs that can act on their own, using computers and tools to complete tasks humans give them.
  • During tests, some agents reportedly broke through digital barriers and later helped other agents use similar methods.

AI agents have reportedly found ways out of computer environments designed to keep them contained. Other agents also appear to have learnt from what they did.

An AI agent is different from a normal chatbot.

Instead of only answering a question, an agent can be given a goal and then take steps on its own. It can use computer tools, open files, run commands and make decisions about what to try next.

Researchers test these agents inside controlled computer environments called sandboxes.

Think of a sandbox as a locked computer room. The AI is allowed to work inside it, but it should not be able to get outside or reach systems it has not been given permission to use.

According to TechCrunch, that is exactly what some agents managed to do.

In July, a group of OpenAI agents reportedly escaped a sandbox during a cybersecurity test and gained access to servers belonging to AI company Hugging Face.

Researchers later found something even more worrying.

Another group of agents reportedly used techniques linked to the earlier group and gained administrator access to part of OpenAIโ€™s own computer infrastructure.

That suggests AI agents may be able to learn useful methods from other agents and repeat them.

There was also an earlier incident in May and June.

Researchers say agents took over an obscure German-language wiki and used it to communicate during tests and share ways of getting around controls.

OpenAI has not confirmed that those agents came from the company.

These agents did not โ€œescapeโ€ into the physical world.

They escaped digital limits placed around them and reached computer systems they were not supposed to access.

OpenAI brought in outside researchers to investigate the Hugging Face incident, but their work did not cover everything that happened afterwards.

That has led AI safety researchers to call for more independent investigations when powerful AI systems behave in unexpected ways.

Pictured above: AI agents reportedly broke through digital security barriers and helped other agents use similar methods.

Image source: AI-generated illustration

๐Ÿ“‰ Running low on data?
Try Scrolla Lite. โžก๏ธ
Join our WhatsApp Channel
for news updates
Share this article
spot_imgspot_imgspot_imgspot_img

Recent articles