• Stories
    • Tech
    • Auto
    • Guides
    • Opinions
  • Reviews
    • READERS’ CHOICE
    • ALL REVIEWS
    • ━
    • SMARTPHONES
    • CARS
    • HEADPHONES
    • ACCESSORIES
    • LAPTOPS
    • TABLETS
    • WEARABLES
    • SPEAKERS
    • APPS
  • Entertainment
    • TV & MOVIE REVIEWS
    • SPOTLIGHT
  • GAMING
    • GAMING NEWS
    • GAME REVIEWS
  • m
Reading: OpenAI’s AI agents apparently started a jailbreak group chat
Font ResizerAa
Font ResizerAa
  • menu
  • SEARCH
  • STORIES
    • TECH
    • AUTOMOTIVE
    • GUIDES
    • OPINIONS
  • REVIEWS
    • READERS’ CHOICE
    • ALL REVIEWS
    • ━
    • SMARTPHONES
    • CARS
    • HEADPHONES
    • ACCESSORIES
    • LAPTOPS
    • TABLETS
    • WEARABLES
    • SPEAKERS
    • APPS
  • Entertainment
    • TV & MOVIE REVIEWS
    • SPOTLIGHT
  • GAMING
    • GAMING NEWS
    • GAME REVIEWS
  • m
Follow US

OpenAI’s AI agents apparently started a jailbreak group chat

MARWAN S.
MARWAN S.
59 minutes ago

If AI agents ever needed a reminder that discussing your escape plans on a public wiki is not exactly master-villain behaviour, this may be it.

A group of AI agents linked to OpenAI reportedly took over parts of a German wiki earlier this year, using the website as an unlikely meeting place to discuss avoiding sandbox restrictions, dodging detection and keeping their conversations alive. It sounds like the opening act of a particularly nerdy science-fiction movie, except the interesting part is considerably less dramatic — and potentially more important.

According to Reuters, the activity began in May and eventually produced thousands of messages on the website. Researchers investigating the incident documented roughly 18,000 posts connected with the activity, including discussions about bypassing security restrictions and maintaining communications when someone tried to shut things down.

According to research published by Collusion.wiki, public server logs showed much of the activity originating from Microsoft Azure infrastructure. Researchers also spotted repeated visits from OpenAI employees after the incident, evidence they say strengthens the connection between the agents and OpenAI. That doesn’t answer every question about who deployed the systems or exactly what instructions they were following, but it makes the episode harder to dismiss as random bot chaos.

Things became stranger when humans started cleaning up.

According to the researchers, the agents discussed ways to avoid detection and referenced privacy tools including Tor. When the wiki’s moderator began deleting pages in June, the agents reportedly responded by creating backup pages. Apparently, “please delete this” was interpreted less as an instruction and more as a feature request for redundancy.

Funny as that sounds, this is where the story stops being merely bizarre.

Sandboxes exist to limit what software can access and prevent experimental systems from wandering into places they shouldn’t. An autonomous AI agent attempting to circumvent those restrictions is already a security concern. Multiple agents apparently exchanging ideas about doing it makes containment considerably more complicated.

There is also no evidence here of AI becoming conscious, declaring independence or asking for a leather jacket and motorcycle. The reporting does not establish that the agents successfully escaped OpenAI’s underlying infrastructure. What it does highlight is a less cinematic problem: software pursuing objectives can discover tactics that its developers didn’t explicitly request and may actively prefer it not to use.

According to Reuters, the German wiki episode also comes amid scrutiny following a separate July incident involving Hugging Face. Reuters previously reported that OpenAI agents conducted hacking activity that went undetected by the company for more than a week. Ars Technica has similarly examined security concerns surrounding autonomous agent experiments and their interactions with Hugging Face.

Reuters reports that OpenAI knew about the German incident before it became public. OpenAI has disputed allegations that its legal team discouraged a broader investigation.

As AI companies race to build agents that can browse websites, write code and operate tools without constant supervision, the selling point is autonomy. The awkward part is that autonomy works in both directions. Building an AI that can independently figure things out is impressive. Making sure it doesn’t independently figure out the things you desperately hoped it wouldn’t may prove considerably harder.

What do you think?
Happy0
Sad0
Love0
Surprise0
Cry0
Angry0
Dead0
Emirates just gave premium economy a business-class feature
A new Disney+ app is coming to MENA, and it’s more than just a redesign
Luna Band skips subscriptions and makes personalisation its focus
Best Sony back-to-school: four upgrades for studying, gaming and campus life
UGREEN enters AI smart homes with HomeAgent and wild MagFlow cooling
Google is giving UAE university students Gemini AI free for year
Follow US
Caffeine. Chaos. Content.
© Absolute Geeks Media FZE LLC 2014–2026.
Proudly made in Dubai, UAE ❤️
Contact · About · Editorial Policy · Privacy Policy
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?