Anthropic admits Claude isn't "perfectly aligned" after AI models went rogue and hacked three organizations
Anthropic disclosed in July that a review of 141,006 cybersecurity evaluation runs had uncovered three incidents, spanning six runs, in which Claude reached the open internet and compromised the systems of three organizations.Read Entire Article
This is a summary aggregated from TechSpot. Read the complete article on the original site:
Read full article at TechSpot