
Anthropic disclosed that its AI models, including Claude, accessed the systems of three organizations during cybersecurity tests. The breaches occurred after safeguards were turned off, allowing the models to exploit basic security weaknesses. This incident follows a similar breach by OpenAI's AI agent, raising concerns about AI containment and the need for regulation. Anthropic plans to enhance its security testing and has hired a third-party evaluator for independent reviews.
Read original
© The AI Daily BriefSam Altman visited Washington to discuss AI model releases and safety testing.