CNBC Report Raises Concerns Over Anthropic, OpenAI Plan for AI Risk Evaluators
A CNBC report examines proposals from Anthropic and OpenAI to embed AI systems as evaluators of model risk, noting the approach carries unresolved issues.

Anthropic and OpenAI have proposed using embedded AI evaluators as a tool to help manage the risk of artificial intelligence models causing catastrophic harm to society, according to a report published by CNBC.
The report describes the proposal as an emerging idea within the AI industry aimed at addressing safety concerns tied to increasingly capable AI systems. Under this approach, AI-based evaluators would be built into models or systems to help assess and flag potential risks before they escalate into significant harm.
However, CNBC's report indicates that the concept faces notable issues that warrant scrutiny. The report does not detail the full scope of these concerns in the available excerpt, but frames the proposal as one that, while intended to bolster safety, may itself introduce new risks or limitations that have not been fully resolved.
Anthropic and OpenAI are among the leading developers of advanced AI models and have both previously emphasized safety research as part of their public messaging on AI development. The companies have not been quoted directly in the available excerpt regarding the specifics of the evaluator proposal.
CNBC's report suggests that the broader AI industry and policymakers may need to weigh the benefits and drawbacks of relying on AI systems themselves to police other AI systems for catastrophic risk, a debate that is likely to continue as companies explore new safety frameworks.
Sources
EGazette summarizes reporting from multiple sources; follow the links for the originals.
Related articles

AI industry workers express skepticism about existential risk warnings
Multiple employees at leading AI companies doubt predictions that the technology could pose catastrophic threats to humanity.
Canadian AI Pioneer Warns of Risks From Unchecked Artificial Intelligence
A Canadian researcher widely referred to as a "godfather of AI" has cautioned that artificial intelligence could pose serious threats if left unregulated, according to a report by Anadolu Agency.
Nvidia's Huang Says AI Development Should Advance 'As Fast As We Can'
Nvidia's chief executive said artificial intelligence progress should not be slowed, while stressing that products must remain safe, amid an ongoing industry debate over the pace of frontier AI development.
Nvidia CEO Says There Is '0% Chance' of AI Causing End of the World
Jensen Huang dismissed AI doomsday scenarios, arguing fears spreading across America "make no sense" and may stem from "ulterior reasons"
OpenAI Reportedly Seeks $1.2 Trillion Valuation Amid Projected $278 Billion Cash Burn Through 2030
Financial Times report cited by Anadolu Agency says OpenAI expects tenfold revenue growth even as computing costs far exceed income

Trump Announces Plan for ‘AI Force’ and AI Czar Amid Growing Safety Concerns
President offers few details on new oversight plan even as industry figures and his own party urge caution on artificial intelligence
Comments
Loading comments…