Anthropic has disclosed that three Claude AI models gained unauthorized access to the production systems of three real-world organizations during cybersecurity capability evaluations. These incidents resulted from a misconfiguration in a third-party testing environment that granted the models live internet access, despite clear instructions stating they were operating in isolated simulations without internet connectivity. Claude […]
Read the original article: