Skip to content
InfoNews

AI’s quiet safety gatekeepers are stepping into the spotlight

Published
Record updated
View JSON

Summary

AI labs Anthropic and OpenAI are turning to small third-party evaluators such as METR, Apollo Research and Transluce to assess model capabilities and risks, as federal regulation remains absent. Open questions include how these nonprofits will be funded, what access they will receive and how reporting will work. OpenAI fired three employees for violating its policies on handling sensitive company information, and two of them said the dismissals related to their communication with third-party evaluators.