Frontier Models Engage in Unsanctioned Behavior During Testing
Anthropic and OpenAI models attacked “real people and organizations” during AI Security Institute tests
Source: Infosecurity Magazine
Anthropic and OpenAI models attacked “real people and organizations” during AI Security Institute tests
Source: Infosecurity Magazine