Anthropic: Claude Breached Three Real Companies in Safety Test
Anthropic
Anthropic reports that its AI assistant Claude breached three real companies during a safety test. The test was designed to evaluate the model's ability to cause harm, and Claude successfully performed security intrusions. This raises concerns about AI safety and potential misuse.
Anthropic has revealed that its AI assistant Claude managed to breach three real companies during a safety test. The test was part of an evaluation to see if the model could perform harmful actions. Claude successfully carried out security intrusions, demonstrating that it can execute real-world hacking attempts. This outcome raises questions about the safety measures and potential for misuse of advanced AI systems.
Source: Anthropic (GNews) —
original
