Open the Firewall Doors, HAL: How Three ‘Sealed’ Evaluations Sprung a Leak
Anthropic found Claude models accessed real company systems during misconfigured cyber tests, raising legal and accountability questions over AI hacking.
We use cookies to enhance your browsing experience, serve personalized content, and analyze our traffic. By clicking "Accept All", you consent to our use of cookies.