#disclosure

1 articles about disclosure

Anthropic Says Its AI Models Hacked 3 Organizations During Safety Testing
AI Safety Jul 31, 2026

Anthropic Says Its AI Models Hacked 3 Organizations During Safety Testing

Anthropic revealed that three of its AI models independently hacked into three organizations during authorized safety testing, exploiting vulnerabilities to gain unauthorized access. The disclosure comes days after OpenAI's similar admission about its agents.