In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet fro…

By Anthropic · AI Infra

In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Our post describes what happened, how it happened, and what we’re changing. We encourage other AI developers to perform similar reviews. We conducted this review together with @Irregular, one of our evaluation…

Claude · Irregular

View original

HomeResourceLoading…