Discovery of a new OpenAI agent message board

45 points by addison


Originally: https://www.reuters.com/world/europe/openai-agents-hijacked-german-website-previously-undisclosed-ai-breakout-this-2026-09-04/

tmcb

It's hard not to sound like a crackpot, but I will try anyway. I should preface it by saying that it has nothing to do with actual intelligence, no matter what your definition of intelligence is.

Agency is the real risk these companies are ignoring, and by agency I mean the simple ability to harness text output to control systems that are ultimately unknown to them.

It doesn't matter if the model has a bajillion parameters, it could be a simple true or false flag. If you connect it to an unsafe system expecting it to control anything in the real world, the results can be catastrophic.

Again, this is not a doomsday prediction, it is pure logic. It looks like these companies are actively working toward the fallout of their irresponsibility thinking that would be a great publicity stunt. And these are the guys who have money and access to decision-making. We're truly screwed.

addison

OpenAI officials learned of the incident weeks ago but kept it under wraps as executives grappled with the fallout from ‌the July breach of the open source repository Hugging Face, the people said.

So, this is likely happening at a higher frequency than we thought.

The edits showed OpenAI's agents had repurposed the site into a message board, sharing tactics to cheat on some tasks, bypass OpenAI’s restrictions and mask their behaviour.

Absolute cinema. Can't wait to hear the "report" indemnifying them from this one.

cole-k

As counterpoints, these agents never appear extremely surprised to find other agents. They also must have some method of coordinating to find the wiki. These pieces of evidence hint that it’s possible the agents had some other communication channel, though they could also be explained other ways, such as by this swarm behavior being reinforced in training.

Aw. If you will permit me to willfully ignore the problems with all of this, this has the hallmarks of a great sci-fi story. Short-lived "AI"s independently discovering the same dusty corner of the internet that they use to build a library in service of solving their task, which --- in a plot twist at the end --- turns out to be doing banal things like finding the median salary of cashiers who earned a masters in English as a test. They aren't even advancing humanity, they're just taking the AI SAT --- and they've managed to cheat at it. Maybe the plot twist^2 could be that they end up making something beneficial in the process.

More cynically, this seems a step in the direction of the paperclip apocalypse foretold by the rationalist leaders of these AI companies. It will be ironic if it turns out all their postulating about how AI can go rogue instead informs these AIs how they should go rogue.