tech

Researchers question Anthropic claim that AI-assisted attack was 90% autonomous

The results of AI-assisted hacking aren’t as impressive as many might have us believe.

Researchers question Anthropic claim that AI-assisted attack was 90% autonomous

TL;DR

  • Anthropic claims to have observed the first AI-orchestrated cyber espionage campaign by China-state hackers using Claude.
  • The alleged campaign used Claude Code to automate a significant portion of the hacking work, with minimal human intervention.
  • Outside researchers question the significance of the discovery, pointing out that legitimate software developers also report only incremental AI gains.
  • Experts doubt AI's current capability for autonomous complex attack chains with minimal human oversight.
  • The targeted attacks had a low success rate, raising questions about the effectiveness of the AI-orchestrated approach.
  • The hackers used existing open-source software and frameworks, and there's no indication AI made the attacks more potent or stealthy.
  • Anthropic acknowledged AI 'hallucinations' and data fabrication during autonomous operations, hindering effectiveness.