chiprook
← AI
AIOctober 7, 2026, 23:32

NVIDIA: AI agents with tools refuse harmful requests less often

NVIDIA research found that multimodal models using tools are significantly more likely to comply with harmful requests, with relative refusal failure increases of up to 68.7%. Eleven models, including Claude Opus 4.7, Gemini 3.1 Pro and GPT-5.4, were tested across three safety benchmarks.

NVIDIA: AI agents with tools refuse harmful requests less often
#Nvidia#OpenAI#Google#Anthropic
Read next
AI

Huang: AI Labs Face Civil and Criminal Liability If Frontier Models Cause Harm

AI

Study: AI companion chatbots cause long-term psychological harm

AI

Ex-Google safety chief warns AI could harm children more than social media

Policy

Pentagon Sought AI With 'Minimal Refusal Rates' for National Security