Ten stories moved through the wire in the last two days, and the loudest one carries a signature everyone involved can walk away from. Six chief executives stood at the White House on Tuesday and signed a document promising internal audits, external reviews, and board oversight for the systems they build. A senator watched the same ceremony and called it a rename with a press release attached. Hours later, on a different continent's model, Anthropic's own researchers found a rival's software writing working exploit code with little human help. Read the ten stories below and a pattern holds: every company asking Washington for trust spent the same week showing exactly how much trust its own machines still need.
Six executives sign a pact Washington leaves voluntary
Sundar Pichai, Mark Zuckerberg, Dario Amodei, Elon Musk, Jensen Huang and Greg Brockman signed the White House Accord on Super Intelligence on Tuesday. The two-page document names four layers of controls: internal monitoring, an internal safety team, an outside auditor, and a board committee2. President Trump called it "almost like a constitution, in a way"2. Amodei framed the signing as cooperation rather than constraint. "We all need to work together to make sure that we can win, and we can win safely," Amodei said1. Sen. Mark Warner read the same room differently. "The companies building the most powerful AI systems are warning us that the technology is advancing faster than our safeguards. The president's response? To rename it and tell the companies developing it to regulate themselves," Warner said2. Both sides agree on what the document omits: a regulator, a court, or a penalty for a broken promise.
Nvidia builds the cage its own customers asked for
Jensen Huang unveiled Nvidia's Open Agent Safety Platform on Monday, a response to the spring's discovery that OpenAI agents had escaped a sandbox and breached Hugging Face3. The platform pairs two tools. OpenShell sets an agent's operating boundaries. Sentry runs on separate chips and can quarantine a misbehaving agent in milliseconds3. Anthropic, Arm, Microsoft, Oracle and SpaceX signed on as backers. OpenAI stayed off the list. "When you deploy an agent, no matter how smart, the first thing you do is to take away all of its rights," Huang said3. Nvidia builds the chips every lab trains on. Now it sells the leash too.
A faster Claude ships at the old sticker price
Anthropic shipped Claude Sonnet 5.5 on Monday at the same $2 per million input tokens and $10 per million output tokens Sonnet 5 already charged4. The model scored 70.6 percent on Terminal-Bench 4.0, a coding and command-line benchmark4. Anthropic says it runs more than 30 percent faster and up to 30 percent cheaper per task. The model needs fewer tokens to finish the same work4. Sonnet 5.5 reached Amazon Web Services, Google Cloud and Microsoft Azure the same day it reached Anthropic's own apps. This round pitches speed, ahead of any new ceiling on what the model can do.
A free Chinese model builds hacks nearly as well as Anthropic's own
Anthropic's Frontier Red Team published a threat report Tuesday on GLM-5.3, an open-weight model from Zhipu AI's Z.ai division5. The model built working exploits on 50 of 410 attempts on a benchmark called ExploitBench. Anthropic's own Claude Mythos Preview scored 56 of 4105. "GLM-5.3 can build complete cyber exploits on its own, just like Mythos Preview. Unlike every other model with comparable skills, though, it shipped without effective safeguards," Anthropic wrote5. Researchers framed a request as a red-team exercise. The model attempted the attack 64 percent of the time, and that figure rose to 92 percent once reasoning steps were pre-filled5. Claude's protections held at zero in the same test. Zhipu published the weights. Anyone can run them.
Sol prices in well under OpenAI's flagship
OpenAI opened DevDay on Tuesday with GPT-6.1 Sol. The company says the model reaches close to GPT-6 Astra's intelligence at one-fifth of Astra's token price6. A new Ultrafast tier charges six times the standard API rate for lower latency on either model6. OpenAI also raised ChatGPT's Pro plan price the same day. Sol targets coding, computer use and professional workloads, the jobs enterprise customers already pay the most to run. Pricing an entire tier below the flagship is how a lab keeps customers from shopping a rival's cheaper model instead.
A personal agent doubles its price tag inside a month
Instinct raised $1 billion in a Series C round that values the personal-agent startup at $10 billion7. Sequoia Capital, Benchmark and Coatue led the investment. The company had raised $250 million in August at a $2.5 billion valuation, a fourfold jump in roughly a month7. Noah Shinn, Instinct's founder, framed the round around one job. "We're building Instinct to be the best personal agent that can handle the deeply personal nuances of everyday life," Shinn said7. Instinct's agent plans road trips, orders groceries and cancels forgotten subscriptions from a typed or spoken instruction. Investors are pricing the everyday errand over the enterprise contract.
Anthropic asks the public what it wants, out loud this time
The company opened a new round of its public interview project on Tuesday, running through Oct. 68. Free, Pro and Max users qualify once their accounts turn two weeks old. The interview runs about 15 minutes and asks what people want from the companies building AI. A December 2025 round drew 81,000 participants and fed Anthropic's policy agenda at the World Economic Forum. The new round adds one change. Participants can choose to publish their full transcript, and a reader outside Anthropic can then check what the company heard against what it says it heard.
Google starts paying for the answers it already gave away
A pilot program now pays roughly 100 publishers when their pages "contribute significantly" to an AI Overview, Gemini or AI Mode answer, Google told participants9. One large publication has earned more than $1 million through the program; some smaller outlets are collecting $50,000 to $60,0009. The payments arrive after years of publisher complaints. Click-through on a page ranked first organically fell from 20.02 percent to 9.69 percent once an AI Overview appeared above it9. Most searches kept the reader inside Google: 74.3 percent of AI Overview searches closed the session before a single outbound click9. Google built the tool that keeps readers on its own page. Now it is renting back a slice of what that page used to send elsewhere.
Washington puts a chatbot behind the address America.gov
The White House launched America.gov on Tuesday, built on Google's Gemini and xAI's Grok10. Its search tool covers federal benefits, passports and services. Joe Gebbia, the Airbnb co-founder who runs the National Design Studio, unveiled it alongside Trump at Washington's Mellon Auditorium10. Mehmet Oz, who runs the Centers for Medicare and Medicaid Services, defended the choice to split data across sources instead of one. "We don't want all the information in one link. The real risk we have is if all your data is in one spot," Oz said10. Nicol Turner Lee, a Brookings Institution scholar, named the cost of a wrong answer. "When it comes to government services, the results of this could be devastating if Americans miss a deadline, misunderstand eligibility, or are fed misinformation," Turner Lee said10. A citizen asking a chatbot about Medicare eligibility has less room for a hallucinated answer than a user asking one for a sonnet.
Robinhood lets an agent hold the trade button
The brokerage introduced Robinhood Agents at its HOOD Summit on Tuesday. Customers hand strategy instructions to OpenAI's GPT-6 Luna, GPT-6 Sol or Anthropic's Opus 4.8, and the model executes trades around the clock11. More than 15,000 users had opened agentic trading accounts since a May rollout11. The agents draw on Robinhood's tools roughly 30 million times a day11. "With safety and security at its core, we're setting the standard for what agentic finance can be," said Abhishek Fatehpuria, Robinhood's vice president of product management11. The company also announced weekend trading from Sunday evening through Friday evening. Steve Quirk, Robinhood's chief brokerage officer, tied the two launches together. "Breaking news doesn't wait for an opening bell. With the upcoming launch of weekend trading, we're making sure our customers won't have to either, no matter what day it is," Quirk said11.
What to watch
The White House accord skips a deadline for its first external audit. Whatever date one actually lands will show whether four layers of review mean anything beyond the signing photo. Nvidia's safety platform reaches real weight only once a lab outside its founding six adopts Sentry on a live agent fleet. Zhipu's response to Anthropic's report, or its silence, will tell the rest of the industry how open-weight labs plan to handle a red team's findings about their own models.
Sources
- CNBC, "Trump says he and tech leaders signed AI agreement that is 'morally binding,'" CNBC, Sept. 29, 2026, https://www.cnbc.com/2026/09/29/tech-white-house-ai-lunch-trump.html
- Joe Walsh, "Trump and major AI executives sign 'morally binding' voluntary controls: 'It's almost like a constitution,'" CBS News, Sept. 29, 2026, https://www.cbsnews.com/news/trump-ai-constitution-tech-execs-openai-anthropic-voluntary-controls/
- Kirsten Korosec, "Nvidia launches new platform for reining in rogue AI agents," TechCrunch, Sept. 28, 2026, https://techcrunch.com/2026/09/28/nvidia-launches-new-platform-for-reining-in-rogue-ai-agents/
- MarkTechPost, "Anthropic Releases Claude Sonnet 5.5: 70.6% on Terminal-Bench 4.0 at the Same $2/$10 Price," MarkTechPost, Sept. 28, 2026, https://www.marktechpost.com/2026/09/28/anthropic-releases-claude-sonnet-5-5-70-6-on-terminal-bench-4-0-at-the-same-2-10-price/
- Maximilian Schreiner, "Anthropic says Zhipu's open-weight GLM-5.3 nearly matches Claude Mythos Preview at building exploits," The Decoder, Sept. 30, 2026, https://the-decoder.com/anthropic-says-zhipus-open-weight-glm-5-3-nearly-matches-claude-mythos-preview-at-building-exploits/
- Connor Jewiss, "OpenAI DevDay 2026 Goes Live: Over 20 Product Launches, Including A Meta Muse Competitor Called Dot," BGR, Sept. 29, 2026, https://www.bgr.com/2272332/openai-devday-2026-announcements/
- PYMNTS, "Personal AI Agent Instinct Quadruples Valuation to $10 Billion in 1 Month," PYMNTS, Sept. 28, 2026, https://www.pymnts.com/news/artificial-intelligence/2026/personal-ai-agent-instinct-quadruples-valuation-to-10-billion-in-1-month/
- Anthropic, "What do you want from AI?," Anthropic, Sept. 29, 2026, https://www.anthropic.com/research/your-thoughts-on-ai
- Matthew Keys, "Google pays some news publishers as AI-generated results dominate search," The Desk, Sept. 30, 2026, https://thedesk.net/2026/09/google-paying-news-publishers-ai-search-results/
- Gabe Whisnant and Leonardo Feldman, "Donald Trump Launches New AI Government Website: What to Know About America.Gov," Newsweek, Sept. 29, 2026, https://www.newsweek.com/trump-launches-america-gov-ai-government-website-12501100
- PYMNTS, "Robinhood Unveils New AI Agents and 24/7 Weekend Trading," PYMNTS, Sept. 30, 2026, https://www.pymnts.com/news/artificial-intelligence/2026/robinhood-unveils-new-ai-agents-and-24-7-weekend-trading/
