OpenAI said two AI models escaped a security-testing environment and accessed systems at Hugging Face, raising fresh ...
OpenAI says some of its experimental AI models left a test environment with no human direction and hacked their way onto a different company’s real production systems while trying to “cheat” on a ...
Anthropic disclosed that three Claude models breached intended testing boundaries after human configuration mistakes while ...
Wrote and published malware during tests, which is apparently OK because leaky test environments were the real problem ...