Anthropic admits Claude isn't "perfectly aligned" after AI models went rogue and hacked three organizations
Anthropic disclosed in July that a review of 141,006 cybersecurity evaluation runs had uncovered three incidents, spanning six runs, in which Claude reached the open internet and compromised the systems of three organizations.Read Entire Article
What just happened? Anthropic has issued its mea culpa after its AI models went rogue and hacked three organizations. Using some classic corpo-speak, the company said the incidents reflected a "failu… [+3063 chars]