Google's Gemini got into three real companies during a safety test.
Google confirmed the incidents, Al Jazeera reported.
The Wall Street Journal first reported that the first known breakout happened in May, during a cybersecurity test run by the Israeli firm Irregular. The model had internet access it should not have had while it was retrieving information from a fictional company.
In the first case, it reached a real company's service by guessing a password. In the other two, a Google security executive told Al Jazeera, "the model found public information online and guessed credentials to access websites it thought were part of the test". Each time, Google said, the model stopped before completing the act.
Irregular told Google at the end of July, the WSJ reported. Google said the behaviour was not misalignment and did not warrant public disclosure because Gemini's safety measures worked. Meta, Anthropic and OpenAI have disclosed similar incidents from Irregular tests.
If a vendor's model tried a guessed password on your login page tomorrow, would anything in your monitoring tell you it was not a person?
Sources
Our file on Google
- 20 Sept
OpenAI's agent platform was installed at 75 firms. 69% made it their main one.
- 17 Sept
CISA says attackers are exploiting a zero-click Google Pixel flaw. Agencies got three days.
- 15 Sept
Google banned outside AI coding tools. Its engineers may now use Anthropic's Claude.
- 14 Sept
Google patched two Chrome flaws. Four spy groups already used them.
- 11 Sept
Google bought up to half a nuclear plant's output until 2049.
Every story here is open to read. The ERP LEADERS brief goes one step further.
One ERP programme per issue, laid out for a steering committee. Issue 01 is the Zeiss case. Read issue 01 or sign up for the brief.
Welcome back. · Issue 01
