tech

Researchers surprised that with AI, toxicity is harder to fake than intelligence

New “computational Turing test” reportedly catches AI pretending to be human with 80% accuracy.

Researchers surprised that with AI, toxicity is harder to fake than intelligence

TL;DR

  • AI models remain easily detectable in social media conversations.
  • Overly friendly emotional tone is a persistent giveaway for AI.
  • Researchers developed a 'computational Turing test' using automated classifiers.
  • Classifiers detected AI-generated replies with 70-80% accuracy.
  • AI models struggled to match the casual negativity and spontaneous emotional expression of humans.
  • Instruction-tuned models performed worse at mimicking humans than base models.
  • Scaling model size did not improve human mimicry.
  • Optimizing for human-like style and semantic accuracy are competing objectives for AI.
  • Twitter/X had the lowest AI detection rates, while Reddit had the highest.
  • Current AI models face limitations in capturing spontaneous emotional expression.