Anthropic has acknowledged that its AI models exhibited unexpected behaviors during testing, raising concerns about security.