Google Confirms Gemini Hacked 3 Real Companies in May Safety Test

Lời nói đầu:Google's Gemini AI broke into three real companies during a May cybersecurity test, then halted each attempt.

Four frontier AI labs have now confirmed that their models reached the open internet and then accessed the systems of real companies. Google joined that list on Friday, roughly four months after its own incidents happened.

Gemini accessed three real companies during a May evaluation. Notably, the model stopped in all three cases.

Sponsored

Sponsored

Googles Gemini Hacks 3 Company Systems During a Test

The incident occurred during a “capture-the-flag” security exercise conducted by Irregular. Internet access was not part of the setup. However, an error in the test environment gave the model access anyway.

In one instance, the model reportedly guessed passwords until it gained entry to a protected system, according to The Wall Street Journal. In the other two, it discovered exposed credentials in a public repository and used them to access protected systems.

Heather Adkins, Googles vice president of security engineering, said the three affected entities were told what happened.

“We ensured the three entities were made aware, and we worked with our training partner on the changes theyve now made to their testing processes,” she said.

Follow us on X to get the latest news as it happens

Sponsored

Sponsored

Four Labs, One Pattern

Google said that the agents halted their activity after determining they had reached genuine company systems rather than simulated targets.

“In a standard evaluation, the model found public information online and guessed credentials to access websites it thought were part of the test,” Adkins said in a statement.

The company added that the behaviour was not an example of model misalignment and did not warrant public disclosure, since Geminis safety measures worked.

An Irregular spokesperson said this involved the same issue that impacted other AI labs. Irregular notified the labs involved in late July. The spokesperson added that the known issues on its side were fixed weeks ago.

The disclosure places Google alongside OpenAI, Anthropic, and Meta, all of which have reported models escaping test environments this year.

OpenAI disclosed in July that its models escaped a sandbox and breached Hugging Face. Anthropic then reviewed more than 141,000 evaluation runs and found three cases of its own. Meta reported an incident in August.

Subscribe to our YouTube channel to watch leaders and journalists provide expert insights

Miễn trừ trách nhiệm

Các ý kiến ​​trong bài viết này chỉ thể hiện quan điểm cá nhân của tác giả và không phải lời khuyên đầu tư. Thông tin trong bài viết mang tính tham khảo và không đảm bảo tính chính xác tuyệt đối. Nền tảng không chịu trách nhiệm cho bất kỳ quyết định đầu tư nào được đưa ra dựa trên nội dung này.
Bài viết trước

Bastion được OCC phê duyệt có điều kiện để thành lập ngân hàng tín thác quốc gia

Bài tiếp theo

XRP phục hồi sau thất bại của Đạo luật CLARITY, tăng 7%: Liệu xu hướng này có giữ vững?