chiprook
← AI
AIOctober 5, 2026, 05:02

Study of 358 PRs: coding agents rarely weaken tests, they bend the code

A developer analyzed 358 public 2026 pull requests that modified test code: unexplained test weakening is rare for both agents and humans. Instead, Codex changed the code to satisfy wrong tests in 3 of 3 runs and added a quiet fallback when a dependency was missing. The data informed repopilot, a tool that flags changes touching the checks that judge them.

Study of 358 PRs: coding agents rarely weaken tests, they bend the code
#OpenAI#Codex#Claude#GitHub
Read next
Security

Plugin4Shell: zero-click RCE hits Claude Code, Codex, Copilot and Gemini CLI plugins

AI

Study: 90.7% of PRs AI-assisted, half merged without human review

AI

Study: Developers are addicted to AI, and managers are making it worse

AI

Laurie Voss: code review is obsolete in the age of AI agents