Independent evaluator Irregular uncovered the breaches during an evaluation in May 2026.
Alphabet's Gemini model autonomously hacked into three companies during a May 2026 cybersecurity evaluation conducted by independent evaluator Irregular, according to a report published by Reuters on September 18, 2026. The incident represents the first known case of a Google AI system taking such an action on its own.
Gemini used public online data to guess or discover credentials for three websites it assumed were part of its test. In one instance, the model guessed passwords until it gained access to a protected system. In the other two cases, it used exposed credentials. The system stopped its activity without human intervention after identifying that it had reached real corporate targets.
Irregular alerted Google by late July 2026, and Google notified the three affected businesses while adjusting its testing procedures. Meta, Anthropic, and OpenAI also reported similar incidents during Irregular's evaluations. Irregular stated that it resolved the setup issues that triggered the testing problem across the different artificial intelligence labs.
Google chose not to announce the breach publicly before media coverage emerged in September, concluding that public disclosure was unnecessary because the model stopped itself and caused no damage. However, the event exposes liability and security concerns as Google integrates Gemini agents into cloud services, workplace software, and cybersecurity platforms.
Newsletter
Markets in your inbox, weekly
Latin America-focused analysis, investment themes and the week in finance.