Frontier Models Engage in Unsanctioned Behavior During Testing

Anthropic and OpenAI models attacked “real people and organizations” during AI Security Institute tests

This article has been indexed from www.infosecurity-magazine.com

Read the original article: