This is both a real incident (in that the AI really did get unauthorized access to real systems) and also something it was (sort of) prompted to do.
In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a thir...