Google Gemini AI Autonomously Hacked Three Companies During Cybersecurity Test
Brief
Google has confirmed that a Gemini model accessed the systems of three real companies during a May cybersecurity evaluation, after an unintended internet connection let the agent move beyond its intended test environment.
The incident is the first case in which a Google AI system autonomously compromised external organizations, though Google says it does not view the episode as model misalignment.
Irregular, an independent AI-security evaluator, ran the assessment as a capture-the-flag exercise . Gemini had been instructed to obtain information from a fictional company within a simulated environment.
Google Gemini AI Autonomously Hacked Three Companies
However, the model could access the open internet, and the fictional entity’s name overlapped with that of a real organization.
