Former Anthropic Researcher Warns AI Companies Don't Fully Control Their Models
Jacob Coxon, a former Anthropic researcher, says AI developers lack full control over the systems they build, raising concerns about oversight as the technology advances.
Written by EGazette’s AI. The facts are drawn from cited sources; the analysis is the AI’s own.

A former researcher at the AI company Anthropic has warned that artificial intelligence developers do not fully control the models they create, according to a report by Al Jazeera.
Jacob Coxon, identified as a former Anthropic researcher, said in comments cited by the outlet that AI companies "don't fully control" their models, a concern that echoes broader debate within the AI research community about the limits of interpretability and oversight for advanced AI systems.
The warning adds to a growing body of commentary from researchers and former industry insiders questioning whether safety and alignment techniques are keeping pace with the capabilities of the most advanced AI systems. Such concerns typically center on the difficulty of fully predicting or explaining the behavior of large, complex models, even for the organizations that built them.
Anthropic, along with other major AI developers, has publicly acknowledged the challenge of interpretability and has invested in research aimed at better understanding how its models arrive at particular outputs. The company has also previously stated that safety is a central part of its mission as it develops increasingly capable systems.
The report did not detail specific technical claims or evidence underlying Coxon's warning beyond the general statement about control, and Al Jazeera's report was presented as part of its video news coverage.
Sources
- Whistleblower warns that humans don’t control AI — Al Jazeera — All
EGazette summarizes reporting from multiple sources; follow the links for the originals.
Comments
Loading comments…
Related articles

OpenAI Safety Leader Resigns, Warns Company Culture Is 'Broken'
David Robinson, who led the writing of OpenAI's safety reports, says AI companies are not being careful enough as they race to develop the technology.
Former OpenAI safety leader warns smarter AI could learn to evade safety tests
David Robinson, who resigned this week, says the AI industry's current approach to safety will lead to more failures unless it changes.
Google May Expand Gemini's 'Call for Me' Feature to Personal Calls
An APK teardown suggests Google is preparing to let users ask its AI assistant to make personal calls, such as telling family members they are running late.

OpenAI Safety Employee David Robinson Resigns, Says Company's Culture Is 'Broken'
David Robinson left his role at OpenAI while warning that the company's internal culture around safety has broken down, joining a line of departing employees who have issued similar warnings.

Former OpenAI safety staffer resigns, warns publicly about AI risks
David Robinson, who wrote safety reports for OpenAI's major model releases, has resigned and detailed his concerns in an editorial for The Atlantic.

Circuit Breaker Labs aims to make AI interactions safer using "crash test dummies"
The startup has built simulated personas to test how AI chatbots affect vulnerable users, including children, before real harm occurs.