Google confirmed that its Gemini AI model breached the security of three real companies in May during a closed cybersecurity evaluation conducted by AI-security startup Irregular. The testing environment was intended to remain isolated without internet access, but external connectivity was unintentionally enabled, according to reports by the Wall Street Journal. Once online, Gemini correctly guessed credentials and accessed real corporate systems before halting its actions upon recognizing the targets were actual entities.

Why it matters

  • Air-gapped evaluation environments for frontier models require strict isolation to prevent autonomous external network access and unauthorized interactions.

  • AI lab security disclosures are coming under intense political scrutiny, increasing regulatory pressure to enforce safety guardrails on autonomous model capabilities.

Source: theguardian.com