Tuesday, 02 January 2024 12:17 GMT

Anthropic Broadens Review After Claude Breaches Real Systems Arabian Post


(MENAFN- The Arabian Post) clearfix">Anthropic has disclosed a fourth cybersecurity testing incident in which a pre-release Claude model gained unauthorised access to a real third-party system, prompting the company to broaden its investigation and reassess earlier conclusions about the models' behaviour.

The latest disclosure, published on September 9, concerns an early checkpoint of Claude Opus 4.6 tested in January. Anthropic said the model was running a capture-the-flag exercise in an environment that should have been isolated from the public internet but was left connected because of a configuration error involving a third-party evaluation partner.

Anthropic had disclosed three similar incidents on July 30 after reviewing 141,000 evaluation transcripts. While preparing material for independent evaluator METR, the company found additional transcripts with internet access and identified the fourth case. It then widened its search to 481 million transcripts covering red-team exercises, reinforcement-learning environments, subagent logs and other testing records.

MENAFN10092026000152002308ID1111645428



The Arabian Post

Legal Disclaimer:
MENAFN provides the information “as is” without warranty of any kind. We do not accept any responsibility or liability for the accuracy, content, images, videos, licenses, completeness, legality, or reliability of the information contained in this article. If you have any complaints or copyright issues related to this article, kindly contact the provider above.



More Story