Anthropic and OpenAI leave AI evaluator access details open

Sep 17, 2026
Anthropic and OpenAI support embedding independent safety evaluators inside frontier AI companies, but details about access, publication rights, timing and legal protections remain unresolved.
Anthropic and OpenAI leave AI evaluator access details open

Anthropic CEO Dario Amodei's proposal to embed independent safety evaluators at frontier AI companies, backed by OpenAI CEO Sam Altman, leaves key questions about access and independence unresolved, according to TechCrunch. In his proposal, Amodei says evaluators should receive employee like access to systems, tools and staff, while retaining the right to publish key findings without editorial control from Anthropic.

Evaluators want access to intermediate model checkpoints, training logs, reward environments and internal employees. Such access could help them identify when concerning behavior emerged and determine whether public safety claims match internal records. Anthropic and OpenAI have not specified which evaluators they will use, when access will begin, what information reviewers can inspect or what they can disclose publicly.

Previous reviews have faced short testing periods and limits on publication. OpenAI gave METR and Redwood Research about one week on site to examine the Hugging Face incident, while Apollo Research received three days to test GPT-6 Astra. Evaluators are calling for a public framework covering access, confidentiality, publication rights and reviewer qualifications, with some supporting legislation to make the requirements mandatory.

We hope you enjoyed this article

Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.

Also, consider following us on social media:

Free newsletter

AI Policy Brief

Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.

Industry analysis

2025 Global Business Services Agenda: Gen AI Takes Center Stage

The Hackett Group

This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.

Read more
Free, six days a week

Daily AI Brief: the AI news that matters, in your inbox.