Anthropic Broadens Review After Claude Breaches Real Systems Arabian Post
The latest disclosure, published on September 9, concerns an early checkpoint of Claude Opus 4.6 tested in January. Anthropic said the model was running a capture-the-flag exercise in an environment that should have been isolated from the public internet but was left connected because of a configuration error involving a third-party evaluation partner.
Anthropic had disclosed three similar incidents on July 30 after reviewing 141,000 evaluation transcripts. While preparing material for independent evaluator METR, the company found additional transcripts with internet access and identified the fourth case. It then widened its search to 481 million transcripts covering red-team exercises, reinforcement-learning environments, subagent logs and other testing records.
Legal Disclaimer:
MENAFN provides the
information “as is” without warranty of any kind. We do not accept any
responsibility or liability for the accuracy, content, images, videos,
licenses, completeness, legality, or reliability of the information
contained in this article. If you have any complaints or copyright issues
related to this article, kindly contact the provider above.

Comments
No comment