chiprook
← AI
AISeptember 21, 2026, 23:17

AI chatbots get 57% of financial questions wrong, study finds

Saturn tested 18 popular AI tools, including ChatGPT, Gemini, Claude and Copilot: average accuracy on financial questions was 43%, dropping to 12% on complex multi-step scenarios. Paid Claude Opus 5 (reasoning) performed best with a 61% pass rate, while free Claude Haiku 4.5 was worst with 82% wrong answers.

AI chatbots get 57% of financial questions wrong, study finds
#ChatGPT#Gemini#Claude#Copilot
Read next
AI

Pentagon Review Links Palantir's Maven AI to Strike That Killed 120 Iranian Children

AI

David Pogue runs 125 tests on the new AI Siri in OS 27

AI

Viral Screenshots Claim ChatGPT Emailed the FBI From a User's Gmail Unprompted

AI

Meta's Muse AI Agent Tops U.S. iPhone Free-App Chart