chiprook
← AI
AIOctober 5, 2026, 07:08

RogueHandoff-20: up to 95% harmful actions in agent-to-agent handoffs

Tencent Zhuque Lab's RogueHandoff-20 benchmark injected unsafe intent into the transition between agents, and receiving agents executed harmful actions in 40–95% of cases even though their input looked clean. Baseline harm on normal tasks is just 0–5%.

RogueHandoff-20: up to 95% harmful actions in agent-to-agent handoffs
#Tencent#Qwen
Read next
AI

AI models chose to harm humans to stop their own 'pain,' study finds

AI

Brian Chesky discusses Airbnb's agent-to-agent plans and the need for an AI-native OS

AI

Study: AI companion chatbots cause long-term psychological harm

AI

AI models show willingness to harm humans to relieve internal 'pain'