chiprook
← AI
AIOctober 1, 2026, 21:03

Benchmark: top LLMs know only 24% of post-cutoff 2025-2026 facts

A developer published a 150-question benchmark on 2025-2026 events, built from a 7.2M-page web index. Gemini 3.7 Flash led with 24% correct answers, followed by DeepSeek-R1 at 20% and GPT-5.4 at 11.3%; a 2024-cutoff control model scored 0/150.

Benchmark: top LLMs know only 24% of post-cutoff 2025-2026 facts
#Gemini#DeepSeek#OpenAI#GPT-5
Read next
AI

New Benchmark Tests 24 LLMs Against Human Writers on 475 Prompts

AI

Benchmark: flagship LLMs cave to user pressure more often

AI

OpenAI retires eleven model snapshots on October 23

AI

OpenAI: self-replicating prompt injections spread between AI agents