Subscribe Sign in

The fix for rogue AI agents could be more AI

TechCrunch
1 min read Rewritten in plain language

Artificial intelligenceAI Safety

Show what we removed Rules applied: A1×2 A3×2 C1 D2×4 D3×13 D4 E3 F2×3 all 30 rules
  • As companies hand off longer and tasks to AI agents, they are running into an oversight problem: Agents can act faster, longer, and at greater volume than humans can realistically review.
  • Relying on AI was necessary for the independent investigation of the OpenAI Hugging Face incident.
  • Outsmarting an AI is not hypothetical, he said, pointing back to the OpenAI incident.
  • Some other startups, like Braintrust, LangChain, and Judgment Labs, have raised hundreds of millions of dollars, while more mature companies like Arize and Galileo have already exited.
  • The tool puts yet another AI between a coding agent and its next action, connecting to agentic tools such as Claude Code and Codex.

5 sentences from our version of the report, chosen to cover it. Nothing here is written; every line is in the article below. How

A number of other startups, like Braintrust, LangChain, and Judgment Labs, have raised hundreds of millions of dollars.

Headline check

There is nothing in this headline a machine can check against the report: no figure, no name and no quotation.

Nothing was measured here, so nothing is claimed. How this is checked

Server grill with blue light library picture
Not from this story. A library photograph of computer server technology, used to illustrate it. Server grill with blue light bigpresh / flickr, CC BY

As companies hand off longer and tasks to AI agents, they are running into an oversight problem: Agents can act faster, longer, and at greater volume than humans can realistically review.

The emerging answer from AI labs and startups is both simple and maddening: Put another AI in the loop.

Relying on AI was necessary for the independent investigation of the OpenAI Hugging Face incident. Redwood Research’s chief scientist, Ryan Greenblatt, one of three auditors, jokingly referred to their efforts as a “slop-vestigation,” noting that the volume of data “made it impossible” to understand what was happening without relying on AI.

Outsmarting an AI is not hypothetical, he said, pointing back to the OpenAI incident.

Y Combinator has funded 106 companies related to AI observability in recent years, as TechCrunch counted. Some other startups, like Braintrust, LangChain, and Judgment Labs, have raised hundreds of millions of dollars, while more mature companies like Arize and Galileo have already exited.

For some AI safety researchers, that has meant turning their research on rogue behavior into tools for the corporate sector.

Apollo Research, a public-benefit corporation that studies AI deception, launched an AI monitor called Watcher in February this year after switching its status from nonprofit to a public-benefit corporation. The tool puts yet another AI between a coding agent and its next action, connecting to agentic tools such as Claude Code and Codex.

Apollo uses multiple layers of AI monitors, Kyle Dai, a member of Apollo’s technical staff, said in a written response to TechCrunch.

Shortened to 1 minute of reading, this version reads 7.7 on the Niral Score.

You are reading our version, not theirs. This is TechCrunch's report shortened to its most important sentences, in plainer words, with verdicts and loaded words taken out. Plain description stays, and so do adjectives that carry a fact, such as "former" or "federal". The reporting, the facts and the quotations are theirs — quotations are never edited — and the indicators beside it measure this version. Hover or tap Adjectives to see every one left in the text.

How this outlet filed it, and how we rewrote it

No other newsroom we read has filed on this event, so there is nothing to compare it with yet.

Outlet Niral ScoreAdjectivesSourcingHappiness
TechCrunchas they published this story 9.8 21 63 50.4
Mundane Readneutralized from TechCrunch 7.9 19 63 50.4

Sign in to react.

Comments

Nothing here yet.

Sign in to comment.

Questions

Readers can ask a question about this story here. Questions and answers are for subscribers. Sign in to read them.

Comments are read before they appear where anything in them needs a person to look. Nothing posted here is ever deleted; a comment taken down keeps its text and the reason, so the decision can be looked at again. How this works