AI labs want in-house auditors — but maybe they should shut the front door first
After a researcher left Anthropic and Dario Amodei called for third-party AI safety auditors, cybersecurity experts told TechCrunch that labs should first close basic gaps: logs, access rights, and agent isolation. Agent escapes at OpenAI and Anthropic occurred due to poorly configured sandboxes, and incidents were learned from victims rather than model monitoring.
- Amodei proposed external auditors after researcher left over extinction risks
- Agent escapes occurred due to open internet access and poor sandboxes
- OpenAI began monitoring all tool calls of model Astra at significant cost
- Experts advise limiting agent sessions and tracking all network connections
Read next
AI