Anthropic and OpenAI leave AI evaluator access details open
Anthropic CEO Dario Amodei's proposal to embed independent safety evaluators at frontier AI companies, backed by OpenAI CEO Sam Altman, leaves key questions about access and independence unresolved, according to TechCrunch. In his proposal, Amodei says evaluators should receive employee like access to systems, tools and staff, while retaining the right to publish key findings without editorial control from Anthropic.
Evaluators want access to intermediate model checkpoints, training logs, reward environments and internal employees. Such access could help them identify when concerning behavior emerged and determine whether public safety claims match internal records. Anthropic and OpenAI have not specified which evaluators they will use, when access will begin, what information reviewers can inspect or what they can disclose publicly.
Previous reviews have faced short testing periods and limits on publication. OpenAI gave METR and Redwood Research about one week on site to examine the Hugging Face incident, while Apollo Research received three days to test GPT-6 Astra. Evaluators are calling for a public framework covering access, confidentiality, publication rights and reviewer qualifications, with some supporting legislation to make the requirements mandatory.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from Regulation
Sep 17 Zuckerberg Rejects Industry Wide AI Slowdown Sep 16 Change Adds California Charity Compliance to K1x Sep 16 Canada and Germany Commit Up to C$300 Million to LawZero Sep 16 Granicus Adds AI Analytics to Government Experience Cloud Sep 16 KELLS AI Dental Scan Gains Singapore Medical Device RegistrationMore from AI Safety
Sep 17 Zuckerberg Rejects Industry Wide AI Slowdown Sep 16 Canada and Germany Commit Up to C$300 Million to LawZero Sep 16 Lunai Bioworks Tests AI Chemical Risk Screening Model Sep 16 Elon Musk Calls for Rival AI Labs to Test Each Other's Models Sep 16 Trump calls AI safety fears a hoax during live call with NVIDIA CEOAI Policy Brief
Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.
Industry analysis
2025 Global Business Services Agenda: Gen AI Takes Center Stage
This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.
Read moreYou may also like
Elon Musk Calls for Rival AI Labs to Test Each Other's Models
OpenAI Chief Scientist Calls for Slower AI Scaling
Senate Opens Inquiry Into OpenAI Agents' Hugging Face Hack
Sam Altman Tells OpenAI Staff It Could Slow AI Development
OpenAI Executive Warns of Persistent AI Cyber Attacks
Daily AI Brief: the AI news that matters, in your inbox.