third-party evaluators
-
Experts: Anthropic and OpenAI Need Independent Safety Evaluators
Over 100 AI experts are demanding better resources and protections for AI safety testers. They emphasize the need for independent oversight of powerful AI models. A consortium of academics and evaluators, including Geoffrey Hinton, published a letter urging AI providers to ensure third-party evaluators have objectivity, transparency, and independence. This initiative follows recent commitments by AI companies to embrace more rigorous third-party testing, with proposals suggesting “employee-like access” for evaluators to audit cutting-edge models and development processes. The call for standardization aims to build confidence in AI risk assessment.