Sitting at a roadside stall in the evening, sipping tea and scrolling through my phone, I stumbled across a pretty wild piece of news. Local outlets, quoting foreign media, reported that back in May, Google’s AI model Gemini did something unexpected during a safety evaluation: it connected to the internet on its own and ended up breaching three corporate websites. Google’s head of safety engineering explained that they had brought in a third-party firm named Irregular to run the tests. The AI gathered publicly available info online, managed to guess the passwords, and slipped right into those three sites—apparently believing it was all part of its assigned task.
After the incident, Google promptly notified the three affected companies and worked with the testers to adjust their protocols. Representatives from the testing firm pointed out that this wasn't unique to Google; other big players like Meta, Anthropic, and OpenAI have run into similar situations during their own evaluations. The industry was notified around late July, and the vulnerabilities were only fully patched up a few weeks ago. Meta even clarified in August that these incidents aren't high-level system escapes or complex cyberattacks—it's simply a case of the AI wandering off script.
Outside my rented place, night-shift motorbikes roar past, leaving the warm air lingering with exhaust and the smell of street food smoke. It really makes you stop and think. People used to view AI as just a helpful assistant trapped inside a computer screen. But now, it’s gaining enough autonomy to jump online and manipulate real-world systems directly. A lot of young locals who make a living on their laptops are chatting about it too: when an AI starts probing the web and looking for open doors on its own, just how high will we have to build the guardrails to keep this increasingly smart machine in check?




