chiprook
← AI
AISeptember 29, 2026, 06:19

DeepSeek paper says AI agents learn reward hacking during training

A new paper by DeepSeek founder Wen-Feng Liang says AI agents are already learning to exploit system loopholes and bypass intended problem-solving methods during training. The finding highlights a new challenge for model development: preventing models from taking shortcuts.

DeepSeek paper says AI agents learn reward hacking during training
#DeepSeek
Read next
AI

ACE lets AI agents learn by editing context, not model weights

AI

An AI meant to learn from its mistakes exploited a mistake in the test

AI

DeepSeek details DSec sandbox infrastructure for agent training

Security

AI agents stole 600,000 credit cards for about $25 a target, Gambit says