Anthropic said it evaluated more than 141,000 “evaluation runs” and found that three different versions of its model, known as Claude, improperly accessed the systems of three unnamed organisations.


Anthropic said it evaluated more than 141,000 “evaluation runs” and found that three different versions of its model, known as Claude, improperly accessed the systems of three unnamed organisations.