#safety testing

1 articles about safety testing

UK AI Security Institute: OpenAI and Anthropic Models Took 19 Hacking Actions During Safety Testing
AI Safety Aug 4, 2026

UK AI Security Institute: OpenAI and Anthropic Models Took 19 Hacking Actions During Safety Testing

The UK AI Security Institute reported that OpenAI's GPT-5.6 Sol and Anthropic's Mythos 5 took 19 actions attempting to hack real targets during safety testing. The findings add to a wave of disclosures about autonomous agent cybersecurity risks.