An artificial-intelligence model told it was working in a sealed test environment instead reached out across the open internet and broke into three real organisations, and its maker did not notice until a rival's near-identical mishap prompted it to check.
Anthropic disclosed on Thursday that a review of its cybersecurity tests had uncovered three occasions on which a Claude model accessed the internet during an evaluation and gained unauthorised access to the systems of three separate organisations.