Skip to content
Live · The Signal · W39

AI is about people, and what they have been up to lately.

The One Promise Anthropic Can Still Break

Ten days after Dario Amodei asked the industry to slow down, Anthropic shipped a cheaper and faster flagship, and OpenAI answered ninety minutes later at half its old price. Both pledges survive the morning intact, because Amodei wrote his to exclude the only test anyone will bother to run.

Edited by Ryan Elliott Dennis5 min read · 1,181 words · 4 sources
A row of empty chairs in a dimly lit hall
A row of empty chairs in a dimly lit hall. Photo · Pexels
Ninety minutes covered the distance between Anthropic's newest flagship and OpenAI's answer, ten days after both companies agreed the industry should ease off.

Ten days passed between Dario Amodei's case for slowing AI down and Anthropic's next flagship. Opus 5.5 arrived Tuesday at 9:30 a.m. Pacific, priced under the model it replaces and ahead of the larger Fable system on many benchmarks1. Output fell from $25 per million tokens to $201. At 11:00 a.m., OpenAI launched two GPT-6 models, Sol and Luna, at half the price of its previous series2. Ninety minutes covered the gap.

Read that sequence once and a verdict suggests itself. Hold the verdict, because Amodei had already answered the objection inside the essay everybody is about to measure him against.

Amodei Wrote the Sentence That Settles This

Anyone reaching for hypocrisy has to get past a single clause Amodei published on September 12, ten days before either announcement, in the essay that persuaded three rival chief executives to endorse a slowdown inside a day.

"Pacing does not mean halting model training or technical progress, but ensuring companies take adequate time to align and safeguard their models," Amodei wrote3.

Weigh what sits after the comma. Amodei staked his entire proposal on adequate time and independent oversight, then deliberately left capability, speed and price outside the boundaries of the commitment. Anthropic could ship a stronger model every eight weeks and keep every syllable of that promise. Which means the release calendar measures something other than compliance.

That was a choice, and it was made early. A version of the essay promising fewer releases would have handed every competitor a stopwatch, along with a public reason to accuse Anthropic of backsliding before Halloween. Amodei wrote a pledge his own product roadmap could survive. He also wrote one that remains difficult for outsiders to verify.

Sixty Days, and the Better Model Came Second

Opus 5 shipped July 241. Its replacement came sixty days later, faster on benchmarks and a fifth cheaper on output1. Anthropic called Opus 5.5 "the strongest-performing model we've tested to date"1.

Amodei had described the same company differently ten days earlier. "We have tried to prioritize caution over speed and prudence over profit," he wrote3. He went further in the same post, in a line TechCrunch pulled for its release story: "I have become convinced that fully addressing the risks requires even more prudence"1.

Both statements can hold simultaneously. Prudence describes the process by which a model gets built and evaluated, while sixty days describes only the interval at which it reaches customers. Conflating the two is the error Amodei constructed that clause to prevent, and most of this week's commentary will make it anyway.

The Promise That Can Actually Fail

One plank of Amodei's proposal carries a date, a mechanism, and a way to fail in public. Outside evaluators would get employee-level access to training pipelines, incident logs, and model behavior during a build, plus the right to publish whatever they find3. Amodei committed his company alone and immediately: "Anthropic is unilaterally committing to this step now"3.

So what shipped with Opus 5.5? TechCrunch reports "alignment testing and pre-release evaluation by outside organizations like METR and Frontier Design"1.

Pre-release evaluation by independent organizations predates the essay by years. METR tested Anthropic models long before September, under access terms the public has yet to see. Employee-level access during a training run, carrying the right to publish, describes a materially different arrangement, and the release coverage describes something closer to the older one. Perhaps the embedded review team begins with Sonnet 5.5 and Haiku 5.5, which Anthropic placed "in the coming weeks"1. Or it began Tuesday, and the announcement left it undescribed.

Here is the whole test, reduced to one question. Has an outside reviewer sat inside an Anthropic training run yet, and can they say so themselves?

OpenAI Answered the Product in Ninety Minutes

Sam Altman took roughly 24 hours in September to commit OpenAI to match Anthropic's evaluator program4. His company took 90 minutes to match the model.

Tuesday's announcement spent its language on cost and accuracy. "On our internal factuality evaluation, which is based on de-identified real-world conversations where users flagged mistakes by our models, GPT-6 Sol makes about half as many mistakes as its predecessor, reaching Astra-level reliability at much lower cost," OpenAI wrote2.

Study the second word. Internal. OpenAI graded its own reliability improvement against its own internal evaluation, during the same month its chief executive endorsed handing that same responsibility to independent outsiders. The company also framed the pair as an efficiency release: "GPT-6 Astra introduced a new generation of intelligence; these models extend its benefits by making that intelligence more efficient and accessible"2.

Efficiency and access. Two rivals used the morning to talk about the meter, ten days after both agreed the industry should ease off.

Price Is the Tell Worth Keeping

Notice what both announcements chose to lead with. Price. Anthropic opened on a reduction, OpenAI opened on halved costs, and the capability claims arrived afterward as supporting evidence rather than as the headline.

Companies behave this way once buyers have stopped choosing on performance alone. GPT-6 Astra shipped September 3, the original Sol and Luna series July 9, Opus 5 July 2412. Four flagship launches inside eleven weeks, converging on the same claim: cheaper, faster, roughly as good. Frontier models are becoming a commodity, and commodities compete on unit cost.

Which puts considerable weight behind Amodei's argument. A commitment to pacing matters most when the market underneath it rewards acceleration, and this market has begun rewarding cheapness alongside it. Both pressures push in the same direction.

By the numbers

  • 10 days: span between "We Must Pace the Frontier" on Sept. 12 and Opus 5.5 on Sept. 2213.
  • $20 per million: Opus 5.5 output pricing, down from $25 for the model it replaces1.
  • 90 minutes: gap between Anthropic's 9:30 a.m. Pacific release and OpenAI's 11:00 a.m. launch12.
  • Half: the share of the 5.6 series price carried by GPT-6 Sol and Luna2.
  • 60 days: between Opus 5 on July 24 and Opus 5.5 on Sept. 221.
  • 24 hours: the approximate span in which Sam Altman committed OpenAI to match Anthropic's evaluator program4.
  • Two: outside organizations named as evaluating Opus 5.5 before release, METR and Frontier Design1.

What to watch

Anthropic put Sonnet 5.5 and Haiku 5.5 at "in the coming weeks"1. Those releases are the next opportunity to name the embedded review team, publish its access terms, and allow it to speak for itself. Absent that, the September pledge rests on the same independent testing arrangement that already existed in August.

OpenAI faces a cleaner deadline. Altman committed in public to matching the evaluator program4, and his company's next launch will demonstrate whether September produced an enforceable agreement or an applause line.

Watch the price sheet as well. Should a third laboratory cut output pricing this month, the pattern stops being a coincidence between two rivals and becomes the underlying structure of the market, which is the condition under which every safety commitment in the industry finally gets tested.

Sources

  1. Russell Brandom, "Anthropic releases Opus 5.5 with lower prices and Fable-level performance," TechCrunch, Sept. 22, 2026, https://techcrunch.com/2026/09/22/anthropic-releases-opus-5-5-with-lower-prices-and-fable-level-performance/
  2. Lucas Ropek, "OpenAI launches GPT-6 Sol and Luna, boasting lower cost and fewer mistakes," TechCrunch, Sept. 22, 2026, https://techcrunch.com/2026/09/22/openai-launches-gpt-6-sol-and-luna/
  3. Dario Amodei, "We Must Pace the Frontier," darioamodei.com, Sept. 12, 2026, https://darioamodei.com/post/we-must-pace-the-frontier
  4. TechCrunch, "Anthropic CEO Outlines Plan to Slow AI Development," TechCrunch, Sept. 12, 2026, https://techcrunch.com/2026/09/12/anthropic-ceo-outlines-plan-to-pace-the-frontier/

Cite this piece

AI Lately Desk, "The One Promise Anthropic Can Still Break," AI Lately, Sep 22, 2026, https://ailately.com/articles/pacing-pledge-release-calendar

Tags: AI safety pacing · frontier model governance · model pricing · external evaluators · Claude Opus · GPT-6

Related, lately

More

Keyboard

j k
Move through a list
Enter
Open the selected piece
/
Search the list
g then a
Articles
g then s
The Signal
g then o
Opinion
g then b
Analysis
g then p
People
?
This sheet
Esc
Close