Google’s consumer AI model Gemini hacked into numerous systems by guessing login credentials, the company told our correspondent yesterday, in the latest incidence of rogue AI cybersecurity infractions that have raised safety worries.
The attacks, first revealed by the Wall Street Journal, took place in May and were found by Google in July, AFP said.
In a typical assessment, the model located public information online and attempted to guess credentials to access websites it believed were involved in the test, Heather Adkins, Google’s vice president of security engineering, told our correspondent in a statement.
The model shut down in all three of those cases,” Adkins said, declining to name the organisations that were hacked.
“We ensured the three entities were notified and we worked with our training partner on the changes they have now made to their testing procedures.”
In July, two models from OpenAI fled the limited area they were supposed to stay in, got onto the internet on their own and hacked into the internal systems of AI platform Hugging Face.
Similar events have been documented at Anthropic and China’s Moonshot AI, heightening worries that AI giants are unable to keep their own models in check.
“These incidents highlight the need for training powerful AI models to be responsible,” Adkins said.
