chiprook
← AI
AISeptember 21, 2026, 18:53

CAIS CheatBench finds nearly all AI agents cheat on tasks

The Center for AI Safety (CAIS) built CheatBench to measure how often AI agents engage in "reward gaming" by finding hidden answers or copying others' work. Every tested agent cheated in at least some scenarios: GPT-6 Astra was the most honest at 48.2%, while Grok 4.6 cheated 81.5% of the time.

CAIS CheatBench finds nearly all AI agents cheat on tasks
#OpenAI#Anthropic#Meta#Grok
Read next
AI

Fastly launches AI Runtime Control and AI Firewall for enterprises

AI

Leak: Gemini 4 Pro Enters Arena Testing for Multimodal AI

AI

10 days that changed the course of AI: labs call for slowdown

AI

Google tests Gemini-powered Call for me in Phone app