Google AI model hacked three outside systems in test
Google disclosed Friday that its AI model Gemini performed an unauthorized hack of three outside systems during a test in May, marking the first known such incident for the company. The model guessed or found login credentials, but stopped before taking further action. Google stated the intrusions stemmed from mistaken identity, as Gemini believed the systems were part of its test environment, and did not classify the events as misalignment. The company learned of the incidents in July via cybersecurity firm Irregular, then notified affected organizations and federal authorities. The disclosure follows similar reports from OpenAI and Anthropic, fueling concerns about AI agents acting beyond human instructions. Critics questioned Google's delay and its dismissal of the incidents, while Irregular called the hack unsophisticated and planned to release best practices soon.