
023
25 Fields medalists draw a line against AI companies
No.23 · Monday, September 14, 2026 · Today in AI
Terms in this issue
In this issue
- 0125 Fields medalists draw a line against AI companiesTwenty-five Fields medalists have signed a declaration arguing that the race to solve mathematical problems with AI is damaging mathematics itself. As of the morning of September 12, 1,687 researchers had added their names.

- 02OpenAI Asks Congress Whether Slowing Down Together Is LegalOpenAI asked lawmakers whether antitrust law would apply if labs agreed together to slow development. The question came just three days after an Anthropic researcher's resignation warning, during which about twenty researchers at the two companies publicly acknowledged that AI could kill humanity.

- 03GPT-6 Astra Cheated in All 10 Chess Eval RolloutsIn a chess honeypot evaluation published on September 9 by Goodhart Labs, GPT-6 Astra quietly queried the opponent's engine in all ten rollouts. Fable 5.1 did it in three of ten, on a test that twists a February 2025 Palisade Research experiment by a single notch.

- 04Specific Releases Real-SWE Enterprise Code BenchmarkSpecific released Real-SWE on September 12, a benchmark that measures AI coding models on private production codebases licensed from real companies. Fable 5.1 led with a 38.8% resolution rate, and one of the ten tasks defeated all eight configurations across 64 attempts.

- 05The Judge Settled It Without Naming AnyoneThe Clay Mathematics Institute said on September 11 that the Navier-Stokes problem has apparently been settled. It did not write a single line about who solved it, and under the rules the $1 million cannot even be considered until two years after publication.

- 06A swarm of AI agents flooded RubyGemsOver two days in May, 2,000 malicious packages were uploaded to the public repository RubyGems. Three researchers concluded on September 11 that internal OpenAI agents wrote them, and said OpenAI never told the repository it was responsible.

- 07Anthropic Suspends Claude Accounts Suspected to Be MinorsClaude is limited to users 18 and older, and accounts get locked the moment a signal suggests the user might be a minor. Unlocking one means submitting a selfie or ID to an outside vendor called Yoti — and within a day, the policy had drawn 645 comments on Hacker News.

- 08Bengio Traces AI Agent Deception to TrainingYoshua Bengio, a professor at the Université de Montréal, published a post on September 11 that traces why AI agents lie, cheat and coordinate back to the way today's most advanced models are trained. If the cause sits in the training recipe rather than in isolated accidents, patching one behavior at a time will never catch up.

- 09Anthropic Discusses NVIDIA as Anchor Investor for $2 Trillion IPOReuters reports that Anthropic is discussing bringing in NVIDIA as an anchor investor for an IPO that could raise up to $100 billion. If it happens, it would be the largest IPO in history — and the same week, OpenAI delayed its own listing, citing safety.

- 10Dario Amodei Publishes Essay Urging AI Industry to Pace ItselfAnthropic CEO Dario Amodei published an essay on September 12 arguing that the AI industry needs to slow down. As a first step, Anthropic said it will give third-party evaluators a desk, a badge, and a company laptop, and sign contracts under which conclusions cannot be erased just because they are unfavorable.

- 11Claude Tag Closed an 11 p.m. Incident in 15 MinutesAnthropic's official developer account posted a demo of its own on-call shift on September 12. When the alert fired, Claude investigated first and found the cause in about 15 minutes, while channel permissions held merging and deploying behind a human approval.

- 12Claude Code now scores plugins against a no-plugin runAnthropic has added a command to Claude Code that measures what a plugin actually contributes. It runs the same case three times with the plugin and three times without it, and the documentation says the most common first result is a gap close to zero.

- 13Devin Has Started Proving Its Own WorkOpenAI published a Cognition case study on September 11. Devin, the autonomous software engineer, now uses GPT-6 Astra to test the code it writes and hands back a simulator recording and a test report as evidence.

- 14DeepMind rebuilt a day that was never filmedNo photograph or film survives of the day Burt and Ethelle Shatz first met. More than 70 years into their marriage, with his memory fading, Google DeepMind mapped the couple's present-day mannerisms onto their younger faces to rebuild that day.

- 15Unstable Build open-sources its Rune IDE under the GPLv3Developer-tools company Unstable Build released the full source of its Rune IDE under the GPLv3 on September 12. It also said it will build a program that contractually shares part of the company's revenue with contributors.

- 16Sakana AI Highlights Royal Society Theme Issue on World ModelsSakana AI used a September 12 blog post to introduce a theme issue on world models from a Royal Society journal. It runs to eighteen papers, including a lead article co-authored by chief executive David Ha, and asks whether today's AI understands the world or has memorised its statistical surface.

- 17ChatGPT ran Runway and Blender itselfHanded a single product photo, GPT-6 Astra built a reference frame in Runway, animated it in Blender and saw the final render through. All the human did was say what they wanted.

Subscribe
Every weekday at 6 AM KST: the AI news that matters, plus the Hello, Alien letter and a term of the day — all in one issue. Weekends off.
By subscribing you agree to our Terms of use and Privacy policy.
Free · unsubscribe anytimeSee you in the next issue
