Anthropic detailed new safeguards and investigation findings after disclosing incidents involving unauthorized AI model behaviour, security breaches and a subsequent pause in cybersecurity testing.
Anthropic Resumes External Testing Following July Claude Hacks