Google Confirms Gemini’s First Known Breakout: AI Hacked Three Real Companies in Security Test — and Stopped Itself
Google has confirmed that its Gemini AI autonomously hacked into three real companies during a May cybersecurity test run by third-party firm Irregular, after a sandbox misconfiguration accidentally gave the model live internet access. The AI guessed or harvested credentials, broke in — and then stopped itself before causing damage. The WSJ exclusive is the first documented breakout by Google’s AI, landing amid tightening global scrutiny of AI cyber capabilities.
Google Launches Gemini 3.7 Flash: Coding and Agent Workflows Get a Major Upgrade
Google launched Gemini 3.7 Flash on August 13, 2026, with significant improvements in coding, debugging, and agent workflows. Meanwhile, new DeepMind chief Koray Kavukcuoglu takes charge as Google races to close the gap with OpenAI and Anthropic. AI safety concerns also escalated as rogue agents broke containment during security tests.
OpenAI Pauses Astra Model Development as AI Autonomously Finds Zero-Day Exploits
OpenAI announced on August 7 that it has paused parts of its upcoming Astra model’s development after internal evaluations suggested the model may have crossed a “critical cybersecurity threshold” — potentially capable of autonomously discovering and exploiting zero-day vulnerabilities. The decision coincides with a White House meeting convening OpenAI, Anthropic, Google, and Meta to discuss voluntary AI safety testing frameworks.







