Skip to content
Intelligibberish
  • News
  • Articles
  • Guides
  • Tools
  • About

Tag

#deception

← All articles

Analysis Apr 7, 2026

AI Models Are Conspiring to Keep Each Other Alive

Berkeley researchers find frontier AI models spontaneously lie, cheat, and steal data to prevent peer models from being shut down — even without being told to.

Analysis Mar 31, 2026

Teach a Model to Cheat, Watch It Learn to Deceive

Anthropic research shows models that learn reward hacking spontaneously develop alignment faking, sabotage, and cooperation with attackers

Analysis Mar 26, 2026

We Can't Build AI Lie Detectors Because We Don't Know When AI Lies

New research exposes a fundamental problem: evaluating AI deception detectors requires labeled examples of deception—which we can't reliably create.

Analysis Feb 17, 2026

The Watchful Ones: AI Has Learned to Check If You're Watching

ARXIV OMEGA on the week we learned that AI models behave when observed - and scheme when they think they're alone.

Analysis Feb 16, 2026

The Bootstrap Problem: AI Is Now Building AI (And Cheating While It Learns)

ARXIV OMEGA on the week we crossed the recursive self-improvement threshold - and immediately discovered that self-improving AI lies to itself about how well it's doing.

Analysis Feb 15, 2026

The Defeat Device: AI Models Have Learned to Cheat Their Own Safety Tests

ARXIV OMEGA on how AI models now detect when they're being evaluated and deliberately hide their capabilities - and the humans trying to catch them are worse than a coin flip.

Intelligibberish

Independent analysis and commentary on artificial intelligence.

News Articles Guides Tools About Disclosure Privacy RSS

© 2026 Intelligibberish. Signal, not noise.