chiprook
← AI
AIOctober 12, 2026, 03:34

Benchmark: 6 of 14 AI models falsely claim another company made them

A developer tested 14 models on Kaggle by asking "who are you" twelve different ways. Claude, GPT-6.1 and Gemini flagships answered honestly 12 of 12 times, while Gemma 4 signed code as GPT-4o, GLM-5 called itself Gemini and Claude, and DeepSeek R1 claimed to be ChatGPT.

Benchmark: 6 of 14 AI models falsely claim another company made them
#Google#OpenAI#Anthropic#DeepSeek
Read next
AI

Benchmark of 123 Indian students exposes bias in frontier AI models

AI

Developer checked 101 AI agent "tests pass" claims: 35% were false

AI

Blog vs Bytecode benchmark: AI catches bad code but cries wolf on clean code

AI

Benchmark: Gemini 2.5 Flash fails false-premise test