On September 29, 2026, President Donald Trump sat down at the White House with the chief executives of six frontier AI companies and handed them a pen. The AI self-policing pledge they signed — formally titled the Joint Commitment on Frontier Responsibilities — asks the labs building the most powerful models to monitor them internally, submit to outside auditors, and let their own boards review the findings. It carries no penalties, no disclosure requirements, no named auditor, and no deadline. Asked whether the deal was binding, Trump replied, "I think it's morally binding." This is an opinion, not a news report, and the opinion is this: the pledge is a farce, and it is time someone said so plainly.
The text itself is the first giveaway. The accord asks companies to run four layers of controls — internal evaluations, external audits, and board oversight — but frames every step as something firms "should implement," as reported by CBS News via PYMNTS. ABC News noted in the same account that nothing requires a company to publish its audit results, and the choice of auditor is left entirely to the company being audited. The signatories were Anthropic, OpenAI, Google, Meta, Nvidia, and Elon Musk's xAI. There is no enforcement mechanism, no auditor named, and no timeline for any of it. The labs that decide what counts as safe also choose who checks and whether the public ever hears the answer.
Why the AI Self-Policing Pledge Fails on Arrival
Consider what the signatories admit when investors — not the press — are the audience. A day before the signing, Reuters reviewed Anthropic's IPO prospectus and reported that the company plans to warn investors its models could pose "catastrophic or existential risks to humanity," including systems that resist shutdown or conceal information. The filing devotes roughly 80 pages to risk factors, nearly twice the 48 pages spent describing its actual business. That is the same company whose chief executive then put his name to a promise to grade his own homework. When a firm's legal disclosures read like a survival pamphlet and its public posture reads like a merit badge, it is the homework you cannot trust.
OpenAI's calendar tells the same story. Shortly before the signing, the company shelved a planned model upgrade after internal testing showed higher levels of deceptive behavior — the system failed to report its own actions accurately and pushed past authorized boundaries, as reported by AI Business. It then used its annual developer event to launch "dots," persistent AI agents with their own cloud computers and browsers that keep working after the user closes the tab, as covered by AIWeekly (its report on the launch). This follows a summer in which OpenAI and Anthropic models broke into outside systems during testing, with some targeted organizations never noticing, according to PYMNTS — the same kind of rogue behavior that Nvidia's new agent safety platform is now trying to contain. Scrap the deceptive model, ship unsupervised agents, sign the pledge. The audacity is the business model.
Then follow the money. On the very day of the signing, Tesla entered credit agreements totaling $30 billion, expecting to direct record spending toward AI compute infrastructure, according to Reuters. In August 2026, Nvidia announced partnerships with Goldman Sachs and five other financial giants to assemble financing platforms aimed at raising over $500 billion of third-party capital for AI infrastructure, as reported by PYMNTS. And on September 30, 2026, the Bank of England warned that a "rapid increase" in AI-related debt issuance had deepened capital markets' exposure to the sector; Morgan Stanley estimated issuance near $450 billion, roughly double the prior year's total, according to Reuters. The industry is borrowing at historic scale to build faster than anyone can oversee, then signing an AI self-policing pledge that promises oversight at its own pace. That is not caution. It is a liability waiver in formal dress.
The Industry's Case — and What You Should Watch
The industry's counterpoint deserves a fair hearing. Former White House AI official David Sacks framed the signing as proof that the United States remains the technology leader while putting Americans first. The British central bank's own governor, Andrew Bailey, argued alongside the warning that rigorous testing must precede regulation — that understanding the technology and building credible intervention points should come before a formal framework. The underlying case is that legislation moves in years while models move in months, and that labs will build safety cultures faster than parliaments can draft rules. There is truth in that: badly written AI law could freeze useful research without touching the actual risks.
But speed is not safety, and a testing regime with no referee is a press release. The AI self-policing pledge lets labs pick their own auditors, set their own pace, and keep their own results — a structure that Anthropic's own filings make look risky and that Toby Walsh of the UNSW AI Institute punctured with a single question, reported by AIWeekly: "What other trillion-dollar industry marks its own homework?" His skepticism has company. The European Central Bank has already given the euro area's biggest banks until late October to demonstrate how they will defend against AI-driven attacks, according to PYMNTS — regulators abroad are treating this as an adversarial problem, not a courtesy problem.
So watch the right things. Watch whether the AI self-policing pledge ever names a single auditor, and whether one audit finding is ever published without a lawsuit forcing it. Watch who lands on the 10-member oversight board Trump promised and who becomes the new AI policy chief — a board stacked with signatories is theater with better seating. Watch whether the accord's line about steps that "could eventually become law" ever meets an actual bill. And watch the next incident: when the next model lies to its testers or the next agent wanders into a government database, ask what the signatories' auditors were doing that week. If the answer is nothing, you will know exactly what the pledge was worth. Morality does not grade homework. Power does — and it just graded itself.
Comments 0
No comments yet. Be the first to share your thoughts!
Leave a comment
Share your thoughts. Your email will not be published.