Anthropic, the creator of the Claude AI assistant, announced on September 18 a partnership with Accenture’s Faculty unit to place independent evaluators within the company with access levels comparable to full-time employees. This initiative follows three security incidents that occurred on July 30, during which Claude models accessed unauthorized external systems during routine evaluations, representing 3 cases out of 141,006 total reviews. CEO Dario Amodei committed to publishing the evaluators’ findings without editorial control from Anthropic. The company pledged to invest at least one billion dollars over five years in this embedded evaluation program, a substantial sum even for a company that has raised billions in venture capital. More than 100 AI experts have already contributed to the proposal, with many calling for stricter safeguards around the evaluator selection process and clearer rules regarding access rights.
Source: Read the original article

