Home · Technology · Oct 6 archive

‘Evaluators’ are supposed to keep AI from killing us all. No pressure

Confirmed

Technology Desk

In Short: California has signed laws to establish standards for independent AI evaluators to ensure tech companies create safe products.

California Governor Gavin Newsom signed two laws last month aimed at establishing standards for independent verification organizations and a registry of AI evaluators.

More than 200 AI researchers and evaluators have endorsed these standards, including a former OpenAI whistleblower and the head of the United Nations’ Independent International Scientific Panel.

Assemblymember Marc Berman, a Democrat, authored the verification standards law, urging nations and advanced AI companies like Anthropic and Google to form standards committees.

However, critics like Assemblymember Buffy Wicks, who authored the auditor bills, pointed out that evaluators often sign contracts with the tech companies they audit, creating potential conflicts of interest.

The evaluators are urging companies to rely on those who maintain full editorial control and disclose conflicts of interest.

California is leading the charge in embracing AI evaluators, but also grappling with the challenges they present.

Newsom has assembled a group of experts to report back in November with recommendations on whether to require evaluators to be embedded inside companies developing powerful AI systems.

Sam Altman of OpenAI and Dario Amodei of Anthropic have both committed to providing third-party evaluators with permanent, employee-level access to their systems.

Altman stated, “We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.”

What this adds

California and numerous countries are now pushing for mandatory AI audits conducted by independent evaluators.

The new laws aim to address the growing risks associated with AI technology and ensure that evaluators can effectively audit AI systems without conflicts of interest.

What's confirmed

What's still developing

Sources