The AI companies were contractually scheduled to provide “risk forecasting, and threat ideation exercises” — predictions of pitfalls and dangers their own products might pose in the future. “Frontier AI labs have advanced risk forecasting tradecraft and a clear vision of the next wave of frontier AI technological developments,” the contracts read. “Frontier AI companies can help DoD gain a better understanding of the risks that might be associated with subsequent near-term frontier AI models or features, to avoid strategic surprise and to mitigate the associated risks.”
Heidy Khlaaf, chief scientist at the AI Now Institute and former safety engineer at OpenAI, told The Intercept this is a dubious claim with dangerous implications. Pointing to recent disclosures by OpenAI and Anthropic that semi-autonomous large language models broke into computer networks owned by other companies during testing, Khlaaf questioned their fitness to help build guardrails in the first place. “This is a very concerning development,” she said.
The public would be better served, she said, with a reliance on independent assessment of the risks these companies present, not the word of the companies themselves. Trusting companies to self-report the risks of their own products constitutes both a conflict of interest, Khlaaf added, and a “subversion of democratic processes when AI labs are allowed to take over the arbitration of risk determinations with life-or-death consequences.”
Read the article here.
Research Areas