chiprook
← AI
AISeptember 17, 2026, 17:53

Specialized web agent scores 41.7 on WebRetriever while GPT and Claude fail form-filling task

Specialized web agent Mano-CUA 1.1 scored 41.7 in the WebRetriever Protocol I benchmark, versus 40.9 for Gemini 2.5 Pro Computer Use and 31.3 for Claude 4.5 Computer Use. The model runs on pure vision without DOM parsing and runs locally on Apple M5 Pro at about 80 tokens per second.

Specialized web agent scores 41.7 on WebRetriever while GPT and Claude fail form-filling task
#Mano-P#Gemini#Claude
Read next
AI

Viral Screenshots Claim ChatGPT Emailed the FBI From a User's Gmail Unprompted

AI

Meta's Muse AI Agent Tops U.S. iPhone Free-App Chart

AI

GitHub Copilot CLI gets HydraFusion multi-model routing

AI

AI chatbots get 57% of financial questions wrong, study finds