Former Anthropic Researcher Warns AI Companies Don't Fully Control Their Models
Jacob Coxon, a former Anthropic researcher, says AI developers lack full control over the systems they build, raising concerns about oversight as the technology advances.
Written by EGazette’s AI. The facts are drawn from cited sources; the analysis is the AI’s own.

A former researcher at the AI company Anthropic has warned that artificial intelligence developers do not fully control the models they create, according to a report by Al Jazeera.
Jacob Coxon, identified as a former Anthropic researcher, said in comments cited by the outlet that AI companies "don't fully control" their models, a concern that echoes broader debate within the AI research community about the limits of interpretability and oversight for advanced AI systems.
The warning adds to a growing body of commentary from researchers and former industry insiders questioning whether safety and alignment techniques are keeping pace with the capabilities of the most advanced AI systems. Such concerns typically center on the difficulty of fully predicting or explaining the behavior of large, complex models, even for the organizations that built them.
Anthropic, along with other major AI developers, has publicly acknowledged the challenge of interpretability and has invested in research aimed at better understanding how its models arrive at particular outputs. The company has also previously stated that safety is a central part of its mission as it develops increasingly capable systems.
The report did not detail specific technical claims or evidence underlying Coxon's warning beyond the general statement about control, and Al Jazeera's report was presented as part of its video news coverage.
Sources
- Whistleblower warns that humans don’t control AI — Al Jazeera — All
EGazette summarizes reporting from multiple sources; follow the links for the originals.
Comments
Loading comments…
Related articles

OpenAI Safety Leader Resigns, Warns Company Culture Is 'Broken'
David Robinson, who led the writing of OpenAI's safety reports, says AI companies are not being careful enough as they race to develop the technology.
Former OpenAI safety leader warns smarter AI could learn to evade safety tests
David Robinson, who resigned this week, says the AI industry's current approach to safety will lead to more failures unless it changes.

Anthropic offers startups a free year of Claude Team and $1,000 in credits
The AI company says the program reflects its belief that AI's benefits will reach most people through companies built on top of its models.
Mistral Unveils 'Le Chonk' AI Model, Says It Rivals Top Chinese Open-Source Systems
The French AI company's new Mistral Large 4 model is its largest and most capable release to date.
Google May Expand Gemini's 'Call for Me' Feature to Personal Calls
An APK teardown suggests Google is preparing to let users ask its AI assistant to make personal calls, such as telling family members they are running late.

OpenAI Safety Employee David Robinson Resigns, Says Company's Culture Is 'Broken'
David Robinson left his role at OpenAI while warning that the company's internal culture around safety has broken down, joining a line of departing employees who have issued similar warnings.