EGEGazette
Live
Houthis Accuse Saudi Arabia of 28 Airstrikes in 24 Hours; U.S. Intelligence Chief Visits CairoWashington Warns Citizens of Unexpected Escalation in Middle EastIran Says Strait of Hormuz Will Stay Closed Until US Meets Its ConditionsTrump faces dual setback as U.S. courts block voting and immigration restrictionsLawsuit Filed Against Trump and His Company Over Paid Early Access Service to His PostsTrump Renews Bid to Restrict Birthright Citizenship Through Curbing "Birth Tourism"Trump Cuts Camp David Vacation Short Amid Middle East Escalation WarningsTrump Cuts Short Vacation, Returns to White House as US Issues "Possible Escalation" Warning in Middle EastU.S. Military Announces Four Killed in Strike on Boat in Caribbean4 killed in US military strike on suspected drug-trafficking vessel in Caribbean SeaReporters From CNN, MS NOW and Politico Denied White House Access After Trump BanTrump Cuts Short Camp David Stay Amid Rising Middle East Tensions
opinionReported

Anthropic, OpenAI Propose 'Neutral' AI Evaluators; CNBC Report Flags Concerns

AI labs suggest embedding independent evaluators to guard against catastrophic risks from advanced models, but a CNBC report says the plan raises questions about independence and effectiveness.

· 2 min read · language: en
Article image
CNBC — Top News

Anthropic and OpenAI, two of the leading developers of advanced artificial intelligence systems, have proposed creating embedded AI evaluators intended to help manage the risk that increasingly powerful models could cause catastrophic harm to society, according to a report published by CNBC on Sept. 16.

The report describes the concept as a form of "neutral" oversight mechanism, in which evaluators would be built into or closely tied to the development process of AI models, rather than functioning as fully external or independent watchdogs.

According to CNBC, the proposal is aimed at addressing growing concerns among researchers, policymakers and the public about the potential for advanced AI systems to cause significant harm, whether through misuse, unintended behavior or other failure modes as the technology becomes more capable.

However, the CNBC report indicates that the idea has drawn scrutiny, with questions raised about whether evaluators embedded within or closely affiliated with the companies developing the technology can provide the kind of independent, unbiased assessment that effective oversight would require.

The report does not specify further technical or structural details of the proposal, nor does it identify specific individuals within Anthropic or OpenAI who put forward the plan. CNBC's coverage frames the initiative within a broader ongoing debate over how the AI industry, regulators and outside experts should collaborate—or maintain separation—when it comes to assessing the safety of frontier AI systems.

Anthropic and OpenAI are both prominent developers of large-scale AI models and have previously voiced public commitments to AI safety research. Neither company's direct statements on the specifics of the evaluator proposal were detailed in the available reporting.

The discussion comes amid continued global attention to AI governance, as governments and industry groups weigh various frameworks for monitoring and regulating advanced AI systems.

Sources

EGazette summarizes reporting from multiple sources; follow the links for the originals.

Also available in: ARFRTRUR

Related articles

Comments

Sign in to join the conversation.

Forgot password?

No account?

Loading comments…