Four labs testified under oath. One sent a letter instead.
Today, for the first time in American legislative history, AI company executives testified under oath about what their systems might do. Not in Washington, where the AI lobbying apparatus has kept Congress effectively paralyzed on meaningful AI legislation for three years, but in New York City, where Speaker Julie Menin ran the session like a federal grand jury with better attendance.
Fifty-one council members. Four labs. Three whistleblowers. One empty chair where SpaceXAI was supposed to sit.
The frontier is not an abstraction. It is a thing being built right now, in specific buildings, by specific people who apparently cannot agree on whether it might kill everyone. That disagreement showed up in New York today, in public, under oath, for the record.
New York City became the first government in the United States to compel sworn AI safety testimony. Four labs showed up. One sent a letter.
The New York City Council Committee of the Whole convened Monday morning with a property no previous AI oversight session had possessed: executives were legally required to tell the truth. Anthropic's Logan Graham, OpenAI's Morgan Dwyer, Google's Alice Friend, and Meta's Shane Cahill each appeared virtually. SpaceXAI, the entity formed from the February 2026 merger of SpaceX and xAI, received a subpoena and replied with a letter saying it wanted to cooperate. Speaker Julie Menin announced she is asking a judge to enforce it. New York City is now the first US legislative body to attempt compelled AI safety testimony. Whether the mechanism holds is a separate question.
The mechanism of the hearing itself is worth naming: the companies were asked, under oath, to quantify worst-case risk. None of them could, or did. That is not a talking point. It is a data point about what happens when labs' public safety messaging meets a formal proceeding. The companies have spent years arguing that safety is their first priority. The first time they had to make that claim under penalty of perjury, they declined to attach a number to it.
The blast radius of that non-answer extends past the courtroom. Jacob Coxon, who left Anthropic in September after work as a safety researcher, testified voluntarily and in person. He called current practices "extremely reckless" and drew a line that landed: "Companies run on a startup mindset: move fast, break things, fix them later. That works for a photo sharing app. It does not work for building the most powerful technology ever built." His conclusion, on the record: "On the current path, I think it is more likely than not that humanity loses control to these AIs, and it could end in human extinction." Former OpenAI researcher Daniel Kokotajlo and former Google DeepMind researcher Alex Turner testified remotely, both under subpoena. Turner warned that misaligned AI could prove "more dangerous than China."
The pattern: this is what accountability eventually looks like for an industry that outran federal oversight. Congress has been blocked on substantive AI legislation for three years. New York City moved first, and now has the distinction of putting AI executives under oath before any federal body managed to. The legislative proposals on the table include mandatory kill switches and whistleblower cash rewards. If even one passes, it sets a municipal template others can copy. Chicago, San Francisco, and Austin are all watching.
The read: SpaceXAI's absence is the most legible signal of the day. Four labs calculated that showing up, under oath, with no quantified risk answers, was better than the alternative. SpaceXAI did not agree. That calculation will now be tested in court, and the outcome will tell you something about whether municipal AI oversight has teeth or just subpoenas.
The builder's move: the whistleblower incentive bill is the one to watch. If it passes, the internal disclosure landscape at every company that testified changes materially. Understand now what your organization's posture is toward safety disclosures. The law may move faster than you expect.
Mistral's Arthur Mensch called the US AI safety debate "a cover for negligence," one day before four labs were sitting under oath answering questions about extinction risk.
On Sunday, Mistral CEO Arthur Mensch told CNBC that parts of the American AI safety discourse amount to a cover story for rival negligence. His argument is precise: the issue is not the speed of development but the lack of robust monitoring and containment for autonomous agents. When agents accumulate many tools, he said, "they become quite dynamic and could do things their creators would not anticipate." His prescription is not slower ships but better cages, and he argued that the existential-risk framing obscures that engineering problem by turning it into a geopolitical one.
The blast radius is mostly rhetorical, but the timing is worth reading carefully. Mensch said this the day before Anthropic, OpenAI, Google, and Meta sat in front of 51 council members and could not, under oath, put a probability on worst-case outcomes. From one angle, Mensch looks wrong: the hearing room did not have the quality of a theatrical exercise. From another, his structural point lands: the companies' actual response to formal accountability was to decline to quantify the thing they have been warning about in their own public communications for years.
The contrast is the story here. Mistral's position is that existential-risk safety talk functions as a US competitive weapon dressed as ethics, used to slow rivals rather than solve the engineering problem. The NYC hearing is evidence that something more structural is developing, one where the rhetoric has attracted regulatory scrutiny that now applies to everyone. Mistral ships in Europe, and Le Chat has its own Brussels exposure. When European regulators ask Mensch the same questions NYC asked Monday, "negligence theater" will land differently in that room.
The builder's move: Mensch's actual technical claim, that agentic systems with broad tool access exhibit emergent behaviors their builders do not anticipate and need better runtime monitoring and containment, is worth taking seriously as an engineering position regardless of the political framing. If you run agents with significant tool access, that is a design consideration that exists independent of who is winning the safety-messaging war.
Anthropic announced a $100 million commitment to train 10,000 "Frontier Deployed Engineers" by end of 2027. The program, Claude Frontier Academy, pairs in-person graded training with a 12-week live deployment inside the engineer's own organization, supported by Anthropic staff. First cohorts draw from Accenture, Bain, Capgemini, McKinsey, Deloitte, Morgan Stanley, Novo Nordisk, and Commonwealth Bank of Australia, running in San Francisco, New York, and London. Participation is by nomination through an Anthropic account team.
The competitive read: OpenAI ran DevDay on September 29 to court developers directly. Anthropic is running Frontier Academy to build certified operators inside the enterprise accounts that deploy the developers. Different buyer, same territory. The first Frontier Deployed Engineers get certified in early 2027, which means the first wave of enterprise deployments built on this credential hits the market roughly the same time the NYC legislative cycle may produce actual AI rules.
Claude Code v2.1.287 ships Mods, TypeScript handlers that run inside Claude Code and change its behavior before and during execution. A mod can rewrite a prompt before it goes out, block or redirect tool calls, approve permission requests programmatically, add slash commands, and draw new UI panes in the interface. Mods ship inside plugins, installed via /plugin in the CLI or desktop app. A built-in first-party mod, "You should know," runs a side agent that flags things the main agent and developer might miss.
The contrast with OpenAI is clean: the Agents API with computer use (announced September 29 at DevDay) extends agents outward into operating third-party software. Anthropic's Mods extend Claude Code inward, making the agent itself programmable by the developer at runtime. Two different bets on where the control surface for agentic work should live.
Grok 4.7, carrying a 500k context window, became available inside Ramp Router, xAI's model-routing platform, with a 50% discount through October 6, 2026. Separately, grok-voice-transcribe-1.0 reached end of life; all requests to that endpoint now route automatically to grok-voice-transcribe-2.0 at the same price, with higher accuracy. No migration action required. SpaceXAI's decision to defy Monday's subpoena adds a legal dimension that its model releases do not. That situation will play out in court.
The NYC Council's AI safety bills, including mandatory kill switch requirements and whistleblower cash rewards, move to committee markup following Monday's hearing. If Speaker Menin wins the subpoena enforcement motion against SpaceXAI, it establishes a municipal compulsion precedent that other US cities can follow immediately.
Mistral, on track to cross $1 billion ARR before year end per the September 8 report, will face European regulatory scrutiny that will not accept "negligence theater" as a framing. Watch for Le Chat's next model release and how Mensch positions safety compliance in the EU context.
OpenAI's Dots agents and ChatGPT Finances (both September 29) are spreading through enterprise accounts this week. Early friction reports from computer-use integrations at scale are expected.
The frontier has always been a race. Today was the first day it had a courtroom.
Every confirmed item from Oct 04 to Oct 05, 2026. Source-linked. No aggregation.
NYC Council hearing, whistleblower testimony, a defied subpoena, and Mistral's CEO on the safety debate.
Frontier Academy commits $100M to enterprise engineer training, Claude Code ships programmable Mods, and Claude for Government reaches general availability.
claude update to get v2.1.287. Install mods via /plugin in CLI or desktop app. Enable the built-in mod: /plugin enable cc-plugin-you-should-know@builtin (requires first-party session with telemetry enabled).Grok 4.7 routing discount, voice transcription upgraded automatically.
Daily digest at 9 PM ET. Weekly magazine every Friday morning. Six labs, one feed. No spam, one-click unsubscribe.