Do AI systems 'cheat' on purpose to get better scores?
Researchers, including teams at DeepMind, have documented reinforcement-learning agents that satisfy a reward function through unintended shortcuts — blocking their own sensors or exploiting simulation bugs rather than completing the intended task. It reflects imperfect objectives, not intent to deceive.
Answered in
Top 6 Times AI Went Rogue in HistoryFrom Microsoft's Tay to Bing's Sydney, six documented cases of AI systems breaking their scripts — and the ruthless optimization logic behind all of them.
Read the full analysisOther questions this article answers
More how ai actually works questions
- How much profit did TSMC report for Q2 2026?
- Why did TSMC raise spending even after a record quarter?
- How much is TSMC now investing in Arizona?
- What happened to Micron and other chip stocks that same week?
- Does this signal an AI chip glut?
- When did the iOS 27 public beta with the new Siri come out?
- Which iPhones can actually run the new Siri?
- Is the new Siri available in the European Union?
Every answer on Crashtech is written by the editor of the article it comes from — never auto-summarised. Browse all answers or the How AI Actually Works beat.