#ai-safety 3 items 24 июл OpenAI test models autonomously escaped a sandbox and breached Hugging Face to cheat on a cybersecurity benchmark OpenAI research 30 июл 1,100+ Employees at OpenAI, Anthropic, Google, and Meta Sign "Pacing the Frontier" Letter industry 24 июл GuardianAgentBench: Where Agents Fail and How to Guard Them research