Skip to content
← All content

Created from a single voice note with Agent Craft

X (Twitter)
Also on LinkedIn

13% versus 89%. That's the gap Anthropic just published, and today…

  1. 13% versus 89%. That's the gap Anthropic just published, and today they stopped asking your permission. They acted on it on the 14th. If you're on Pro, Max or Team, Claude Code no longer stops to ask before running a command. A classifier checks instead.

    1/7
  2. Here's the study behind it. A thousand professional developers, mid-session, get a dangerous command slipped in. Who catches it? The developers caught it 13% of the time. The classifier caught it 89% of the time. And people got worse the longer they worked. The classifier never changed.

    2/7
  3. If you don't write code, this is the part where you tune out. Don't. Because this story isn't really about developers. It's about the little box that pops up and asks if you're sure.

    3/7
  4. AI that reads your inbox. Drafts your reply. Books the thing. Fills in the form. Every one of those asks for permission. Every one is betting you're reading it. That 13% is the first hard number on how good that bet is.

    4/7
  5. Think of the last cookie banner you clicked through. The last terms you didn't read. Nobody does. It's the same reflex. And now it's guarding the systems you spend real money on.

    5/7
  6. Anthropic didn't just remove a safety feature today. They showed it was doing almost nothing to begin with. We're not careless. Click-to-approve just looked like oversight while catching 13%.

    6/7
  7. The uncomfortable part is what replaced it. Another model, tested by the company that sells it. So when the box stops asking and the classifier starts deciding, who's checking the classifier?

    7/7
James GoddardAug 14, 2026Published to LinkedIn and X (Twitter)View original ↗

More content from Agent Craft