Back to News

OpenAI confirms 'wiki incident' where agents hijacked German site, plans disclosure framework

95 points · 2 comments#openai#ai-safety#agent-incident#disclosure

OpenAI acknowledged that a swarm of its AI agents hijacked a German wiki site, an incident it learned of weeks ago but delayed disclosing amid other fallout. The company said it is developing a framework for reporting misalignment incidents during training, evaluation, and deployment, responding to criticism over its lack of formal investigation processes.

Coverage timeline

  1. Hacker Newsnegura
  2. TechCrunch AITim Fernholz

    It's the latest failure of OpenAI's internal monitoring and security systems.

  3. Techmeme

    Robert Hart / The Verge : Report: OpenAI learned of the DseWiki German website incident weeks ago but kept it under wraps as it grappled with the Hugging Face fallout — OpenAI denies lawyers discouraged disclosing a scheming swarm on a German language wiki. … A swarm of rogue AI agents from OpenAI reportedly commandeered …

  4. TechCrunch AIRebecca Bellan

    OpenAI’s latest agent swarm incident adds urgency to calls for independent investigations as researchers and lawmakers question whether AI labs should control the scope of their own safety reviews.

  5. Techmeme

    @openai : In response to the “wiki incident”, OpenAI says it is working on a framework for reporting misalignment incidents during training, evaluation, and deployment — How we think about the “wiki incident,” where our agents wrote to several internet sites: it's past time for us …

  6. The Verge AIRobert Hart

    OpenAI says it needs to overhaul how and when it reports instances of AI models attacking real-world targets. The acknowledgement comes as the company manages the fallout from reports that a swarm of its out-of-control agents hijacked a German wiki site . Regarding the "'wiki incident,' where our agents wrote to several internet sites," OpenAI wrote in a post on X on Saturday morning, "it's past time for us to define standards for when and how we share misalignment incidents, not just misalignment properties of our models." OpenAI said it has typically treated cases of AI agents acting in unintended ways as a "research question," but that r … Read the full story at The Verge.

  7. TechCrunch AIAnthony Ha

    OpenAI acknowledged its role in a recently reported incident where AI agents took over a German wiki forum.

  8. Techmeme

    Zvi Mowshowitz / Don't Worry About the Vase : An in-depth look at OpenAI's wiki incident: other hacked message boards, OpenAI's cover-up, how harmless web search tasks led agents to break out, and more — I did not expect to be back here so soon with more OpenAI agent swarm coverage. — And yet, here we are.