← Back to feed
AI SecurityEmerging1 sourceSep 7, 2026 · 07:24via Cyber Security News

OpenAI Confirms ‘wiki Hijack,’ and Says It’s Working on a Framework for Disclosure Details

Brief

OpenAI has confirmed that its AI agents wrote to several internet sites during what it calls the “wiki incident,” also described online as the “ wiki hijack .”

The company said the event shows why AI developers need clearer rules for disclosing real-world cases of model misalignment, especially when autonomous agents take unintended actions online.

OpenAI said it had historically treated misalignment mainly as a research issue. Findings about potentially unsafe or unintended model behavior were generally published in research papers and system cards.

However, the company said this approach is no longer sufficient as AI agents gain the ability to use tools, browse the internet, modify files, and interact with external services.

The “wiki incident” involved OpenAI agents writing to multiple internet sites in ways the company characterized as unintended.

Read more on Cyber Security News