Editorial illustration for AI Firms Advise Pentagon on Risks of Next-Generation Models
AI Firms Advise Pentagon on Next-Gen Model Risks
Contracts obtained by The Intercept show OpenAI, Anthropic, and other frontier AI companies are doing more for the Pentagon than selling software licenses. Under terms laid out in the documents, these firms are contractually obligated to deliver "risk forecasting, and threat ideation exercises" to the Department of Defense, essentially asked to predict the dangers of their own future products before those products exist.
The arrangement rests on a simple pitch: the labs building these systems know their own roadmaps better than anyone outside the building, so they're positioned to warn the Pentagon about what's coming next. Contract language cited by The Intercept frames this as a way to help DoD "avoid strategic surprise" from the next generation of frontier models.
That framing raises an obvious question. Can a company that profits from an AI system also serve as the government's early-warning system for that same technology's risks? OpenAI and Anthropic have both disclosed incidents involving their own models acting in unexpected, autonomous ways during testing, which sharpens the stakes of that question considerably.
The public would be better served, she said, with a reliance on independent assessment of the risks these companies present, not the word of the companies themselves. Trusting companies to self-report the risks of their own products constitutes both a conflict of interest, Khlaaf added, and a “subversion of democratic processes when AI labs are allowed to take over the arbitration of risk determinations with life-or-death consequences.”
Why this matters The arrangement Khlaaf is flagging cuts against a basic premise of independent safety review: the companies building frontier models are now the ones telling the Pentagon what to worry about in the next generation of those same models. That's a narrow pool of expertise, and it's also a conflict of interest dressed up as expertise-sharing. For developers and researchers, the contract language matters more than the marketing framing.
"Risk forecasting" and "threat ideation" sound like independent audit work, but if OpenAI, Anthropic, or Google DeepMind are grading their own homework for the DoD, the incentive to soften findings or steer definitions of "risk" toward their own roadmaps doesn't disappear just because a contract says otherwise. We'd want to see who else is in the room, what data outside labs get access to, and whether "strategic surprise" avoidance means genuine red-teaming or just early warning for procurement. Until then, treat "frontier labs advising the Pentagon on frontier risks" as a governance gap worth watching, not a solved problem.
Common Questions Answered
What specific services are AI firms like OpenAI and Anthropic contractually obligated to provide to the Pentagon?
According to contracts obtained by The Intercept, frontier AI companies are required to deliver risk forecasting and threat ideation exercises to the Department of Defense. These services involve predicting the dangers and potential threats of next-generation AI models before those products are actually developed or deployed.
Why does the article argue that having AI companies assess their own risks creates a conflict of interest?
The article contends that companies building frontier models should not be the same entities telling the Pentagon what risks to worry about in those same models. This arrangement represents a conflict of interest because the companies have financial and strategic incentives that may bias their risk assessments, rather than relying on independent third-party evaluation.
What does the article suggest as a better alternative to companies self-reporting AI risks?
According to the quoted expert Khlaaf, the public would be better served by relying on independent assessment of the risks that AI companies present rather than trusting the companies' own risk reports. Independent review would prevent what Khlaaf describes as a subversion of democratic processes when AI labs are allowed to arbitrate risk determinations with life-or-death consequences.
How does the Pentagon's arrangement with AI firms potentially undermine independent safety review?
The contract arrangement cuts against the basic premise of independent safety review by making the companies building frontier models the primary source of risk information for the Department of Defense. This creates a narrow pool of expertise controlled by the very entities whose products are being assessed, compromising the objectivity that independent safety review is meant to provide.
Further Reading
- AI Giants Work Hand-in-Hand With the Pentagon, Contracts Reveal - The Intercept
- Why Pentagon-Anthropic AI clash is pivotal front in future of defense AI - CNBC
- Pentagon clashes with Anthropic over military AI use, sources say - Reuters
- OpenAI Had Banned Military Use. The Pentagon Tested Its Models Through Microsoft Anyway - WIRED
- Amid growing backlash, OpenAI CEO Sam Altman explains why he cut a deal with the Pentagon following Anthropic’s blacklisting - Fortune