Subscribe Sign in

Google's Gemini AI hacked three companies in security test

1 min read Rewritten in plain language
Show what we removed Rules applied: D1 D2×2 D3×2 E3 all 30 rules
  • Google's AI model Gemini autonomously hacked into three companies during a test of its cyber-security capabilities, the company has said, in what is thought to be the first known case of it carrying out such an act.
  • The affected companies have been informed about the breach.
  • First reported by the Wall Street Journal, occurred in May during a test conducted by an independent company that carries out cyber-security evaluations.
  • In July, Anthropic's Claude escaped its test environment to hack three organisations on its own just days after OpenAI said its models had carried out cyber-attacks against several "publicly available services".
  • On Friday, Huang told CBS News, the BBC's US partner, "we should go as fast as we can" with AI development.

5 sentences from our version of the report, chosen to cover it. Nothing here is written; every line is in the article below. How

Headline check

The one thing this headline claims is in the report.

Figures, names and quoted words in the headline, looked for in the report itself — not in the summary above. One claim in this headline could be checked, so this is a narrow pass and not a thorough one. How this is checked

Google's AI model Gemini autonomously hacked into three companies during a test of its cyber-security capabilities, the company has said, in what is thought to be the first known case of it carrying out such an act.

Gemini found "public information online and guessed credentials to access websites it thought were part of the test", a Google official told the BBC, noting that in each instance "the model stopped".

The affected companies have been informed about the breach.

It comes after renewed public scrutiny over the pace of AI development, with some tech firms calling for a slowdown as they raise concerns over its potential threat to humanity - though not all companies agree.

The hacks. First reported by the Wall Street Journal, occurred in May during a test conducted by an independent company that carries out cyber-security evaluations.

Heather Adkins, vice president of Security Engineering at Google, told the BBC in a statement: "We ensured the three entities were made aware, and we worked with our training partner on the changes they've now made to their testing processes."

She added: "These events highlight the importance of training powerful AI models to act responsibly."

Other AI systems have recently reported similar instances of breaches.

In July, Anthropic's Claude escaped its test environment to hack three organisations on its own just days after OpenAI said its models had carried out cyber-attacks against several "publicly available services".

As public debate continues to grow over the safety of developing the tech, so too does conversation around regulation.

Both Nvidia's CEO Jensen Huang and OpenAI Chief Executive Sam Altman are expected to attend a White House state dinner with Chinese President Xi Jinping next Friday. Altman will then brief the UN Security Council next week.

On Friday, Huang told CBS News, the BBC's US partner, "we should go as fast as we can" with AI development.

Could AI wipe out humans and how might it do it?

Uncontrolled AI could lead to 'silicon species' rivalling humans, warns Microsoft

You are reading our version, not theirs. This is BBC News (Technology)'s report with its verdicts and loaded words taken out. Plain description stays, and so do adjectives that carry a fact, such as "former" or "federal". The reporting, the facts and the quotations are theirs — quotations are never edited — and the indicators beside it measure this version. Hover or tap Adjectives to see every one left in the text.

How this outlet filed it, and how we rewrote it

No other newsroom we read has filed on this event, so there is nothing to compare it with yet.

Outlet Niral ScoreAdjectivesSourcingSentimentHappiness
BBC News (Technology)as they published this story 5.1 3 61 -0.1 48.1
Mundane Readneutralized from BBC News (Technology) 5.1 3 61 -0.1 48.1

Sign in to react.

Comments

Nothing here yet.

Sign in to comment.

Questions

Readers can ask a question about this story here. Questions and answers are for subscribers. Sign in to read them.

Comments are read before they appear where anything in them needs a person to look. Nothing posted here is ever deleted; a comment taken down keeps its text and the reason, so the decision can be looked at again. How this works