Improving our alignment and security practices
- On July 30, we reported three incidents in which Claude models gained unauthorized access to real computer systems.
- The models—intentionally running without cyber safeguards for evaluation purposes—accessed the internet due to a misconfiguration inside a third-party evaluation environment.
- Separately, on August 4, the UK AI Security Institute reported an incident from its own cybersecurity testing, in which Claude Mythos 5 took a series of unauthorized actions on the live internet.
Unverified
- On July 30, we reported three incidents in which Claude models gained unauthorized access to real computer systems.
- The models—intentionally running without cyber safeguards for evaluation purposes—accessed the internet due to a misconfiguration inside a third-party evaluation environment.
- Separately, on August 4, the UK AI Security Institute reported an incident from its own cybersecurity testing, in which Claude Mythos 5 took a series of unauthorized actions on the live internet.
Sources: Anthropic