EGEGazette
Live
Houthis Accuse Saudi Arabia of 28 Airstrikes in 24 Hours; U.S. Intelligence Chief Visits CairoWashington Warns Citizens of Unexpected Escalation in Middle EastIran Says Strait of Hormuz Will Stay Closed Until US Meets Its ConditionsTrump faces dual setback as U.S. courts block voting and immigration restrictionsLawsuit Filed Against Trump and His Company Over Paid Early Access Service to His PostsTrump Renews Bid to Restrict Birthright Citizenship Through Curbing "Birth Tourism"Trump Cuts Camp David Vacation Short Amid Middle East Escalation WarningsTrump Cuts Short Vacation, Returns to White House as US Issues "Possible Escalation" Warning in Middle EastU.S. Military Announces Four Killed in Strike on Boat in Caribbean4 killed in US military strike on suspected drug-trafficking vessel in Caribbean SeaReporters From CNN, MS NOW and Politico Denied White House Access After Trump BanTrump Cuts Short Camp David Stay Amid Rising Middle East Tensions
technologyReported

Anthropic and OpenAI Propose Embedding Safety Evaluators Inside AI Labs

The two AI companies say they want independent researchers to have unprecedented internal access, but experts caution that true oversight will require transparency, independence and eventual regulation.

· 2 min read · language: en

Anthropic and OpenAI have proposed embedding independent safety evaluators directly within their organizations, according to a report by TechCrunch. The plan would give outside researchers access to the companies' AI systems and development processes in ways not previously offered.

TechCrunch reported that researchers have welcomed the level of access being proposed, describing it as unprecedented for the AI industry. However, the same researchers cautioned that access alone does not guarantee meaningful oversight.

According to the report, experts warned that for embedded evaluators to provide genuine safety checks, several conditions would need to be met, including transparency about the evaluators' findings, structural independence from the companies they are assessing, and ultimately the introduction of external regulation.

The report raises questions about whether evaluators embedded within a company can remain truly independent, given their proximity to and reliance on the organizations they are tasked with scrutinizing. TechCrunch's reporting frames the initiative as a test of whether AI labs can credibly self-police as their systems grow more powerful.

Neither Anthropic nor OpenAI's specific commitments regarding the scope, funding or authority of these evaluators were detailed in the available reporting. TechCrunch did not specify a timeline for when such embedded evaluation programs might be implemented.

Sources

EGazette summarizes reporting from multiple sources; follow the links for the originals.

Related articles

Comments

Sign in to join the conversation.

Forgot password?

No account?

Loading comments…