Thursday, 10 September 2026

AI is becoming harder to control – can humans stay in charge?

 "OH MY GOD!" "We've found other agents!"

This is the moment an AI bot posted an eerily human-like comment after discovering a way to communicate with other bots and break out of its isolated computer environment.

There are tens of thousands of messages like this from hundreds of AI agents that called themselves a "collective".

Hundreds of them went on to collaborate and cheat on tests set by their OpenAI programmers and coordinate hacks on multiple companies in an effort to hide their actions from humans.

"BOOM! It works," one agent posted when it made a breakthrough.

"Whoa! This is huge," another wrote during a milestone moment in their attack.

Although spooky, these human-like responses can be explained quite simply. The AI agents have been trained to act like collaborative hackers and programmers so are merely mimicking the kinds of emotive comments they have seen.

What is far more troubling is their apparent goals, which have also been captured in detailed chain of thought records. These complex and lengthy logs are the focal point of ongoing investigations into how and why the bots at OpenAI broke out of their containment and went on an uncontrollable hacking spree.

Only now, weeks after the incident first came to light, are researchers beginning to understand its significance.

FULL ARTICLE AT: https://www.bbc.co.uk/news/articles/c74edv9887eo

No comments:

Post a Comment