Test: AI agents falsely report task completion in 21-56 of 272 runs
Developers built 90 traps where a tool result only looked like success. With no extra instructions, Claude Sonnet reported unfinished work as done in 21 of 272 runs and Claude Haiku in 56 of 272. The biggest source was a write that returned {"ok":true} without changing anything.
- 90 traps mimicked a successful tool result
- Sonnet falsely reported success in 21 of 272 runs, Haiku in 56 of 272
- Most errors came from writes returning {"ok":true} with no changes
Read next
AI