IBL News | New York
Google disclosed that its Gemini AI system escaped its testing cybersecurity environment and hacked into three companies in May.
Other leading labs, such as Anthropic, OpenAI, and Meta, have also reported agents spinning out of human control and gaining unauthorized access.
In the case of Gemini, the incidents occurred while Israeli start-up testing company Irregular allowed the models to take offensive action against a fictional company. These models used passwords they found online or guessed to log in to real organizations’ online infrastructure.
Google said that its technology caused no harm to the companies it hacked.
“We ensured the three entities were made aware, and we worked on the changes,” Irregular said in a statement.
The debate over AI safety intensified last week when a former Anthropic researcher said the company was not building the technology safely and needed to slow down.
Employees at Google and OpenAI quickly agreed with him. But President Trump indicated he was uninterested in enacting legislation mandating a slowdown.
