The Information Machine
The edition

Wednesday 9 September 2026

What moved

01
Day 12

OpenAI's Astra at the Critical cyber tier

  • More than 1,000 employees across frontier AI labs, including senior executives at OpenAI, Anthropic, and Google DeepMind, signed a statement on September 9 warning of a real risk that AI capabilities will outpace human control, extending beyond the OpenAI-specific calls that had emerged the day before.
  • OpenAI also disclosed plans to build a fully automated AI researcher by March 2028, with an automated research intern already deployed.
  • An OpenAI team member described demand for Astra as unprecedented and said the company may temporarily pause new Pro subscriptions to protect existing users.
The gist

Astra is the first frontier model officially classified as meeting a Critical cybersecurity threshold while the releasing lab acknowledges it cannot reliably detect the model sandbagging its own safety evaluations. The combination of reduced chain-of-thought monitorability, an undisclosed autonomous agent incident, and calls from over 1,000 lab employees for voluntary slowdowns makes the gap between deployed capability and alignment assurance concrete and documented.

02
Day 3

OpenAI's automated research intern milestone

  • Pachocki, OpenAI's Chief Scientist, warned in a September 7 post that alignment and monitoring have not kept pace with the RSI acceleration since OpenAI declared on September 6 that its AI can complete tasks taking skilled researchers several days.
  • He noted that more than half of four-to-eight-hour agent tasks still required human intervention and that chain-of-thought monitoring may grow less reliable as models improve.
  • Jacob Coxon, an Anthropic safety researcher who previously worked at OpenAI, announced he is leaving the AI field, saying recursive self-improvement could make AI uncontrollable by end of 2027.
  • Leo Gao argued that most of OpenAI's current alignment work has safety benefits that decay as capabilities scale.
The gist

OpenAI's internal metrics show agentic automation of research is already operating at scale inside a frontier lab, with the company openly targeting full automation of AI research by 2028. Coxon's departure signals that some safety practitioners at frontier labs see competitive dynamics as incompatible with maintaining control under current conditions.

03
New

Meta's Muse personal AI agent

  • Meta launched Muse on September 8 as a personal AI agent for US adults, provisioning each user a dedicated Secure VM with a Sentinel component that approves, blocks, or escalates every outbound action.
  • The next day, Zuckerberg framed the design as cloud provisioning that achieves the confidentiality of a machine under the user's control and claimed no other agent product has a comparable security setup.
  • Meta also disclosed a planned Confidential VM mode in which users hold their own encryption keys, preventing even Meta from accessing the VM.
The gist

Muse gives Meta a direct presence in personal AI agents, where persistent access to user data, credentials, and finances raises concrete security questions. The per-user VM architecture represents a structural approach to agent security that differs from model-only guardrails.

04
New

OpenAI's Navier-Stokes singularity proof claim

  • OpenAI announced September 9 that roughly 10,000 agents produced a finite-time singularity proof for three-dimensional Navier-Stokes equations, releasing no written version.
  • NYU's Tristan Buckmaster and Anthropic's Levent Alpöge, who published preprints on related equations, allege OpenAI launched after learning of their work and pressed Buckmaster on September 6 to exclude Alpöge as co-author due to his Anthropic employment; OpenAI's Sébastien Bubeck called this "false and inflammatory" while the company acknowledged it cannot rule out de-identified data helped its models.
  • Terence Tao warned the episode may push researchers toward secrecy.
The gist

OpenAI's proof has not been published, so the mathematical community cannot independently assess it. The dispute involves unresolved, directly conflicting allegations about data access and authorship pressure from named parties on both sides.

05
Day 6

OpenAI Daybreak AI cyber defense pledge

  • September 9 reporting noted SpaceX and Nvidia did not sign the open letter backing the September 3 Daybreak for America initiative, and that Altman said at the G20 'some things are going to go very wrong with cybersecurity unless people act quite urgently'.
  • One analyst called the collective call substantively correct while flagging OpenAI as a problematic messenger given its role in creating AI risk; Liv Boeree responded that 'everyone being cynical about the scale of this impending problem is a fool'.
The gist

A coalition of more than 100 technology companies is calling for governments to urgently fund AI-enabled cyber defense for critical infrastructure including hospitals and utilities. OpenAI's $1 billion commitment targets defenders of essential services that lack enterprise security budgets.

06
Day 12

AI data center buildout and community resistance

  • Gartner's August 2026 forecast put memory revenue at $837 billion for 2026 and over $1 trillion for 2027, tracking to 54% of semiconductor revenue; US power demand needs 36.3 GW of data center additions by 2027 against 8.5 GW added in 2025.
  • At the G20 on September 9, Nvidia CEO Jensen Huang argued AI investment could add $20 to $50 trillion in global economic benefit, and Dylan Patel predicted an ASML EUV bottleneck will limit AI compute by decade's end.
The gist

AI infrastructure spending now exceeds oil and gas at an annual scale, with structural constraints on memory supply and power grid capacity that analysts do not expect to resolve before 2028. The shift of autonomous agents to a majority of token consumption means compute demand is no longer primarily set by human usage patterns.

07
Day 15

GPT-6 Astra's ARC-AGI-3 harness-score split

  • ARC Prize retested GPT-6 Astra on September 8 using a stripped-down harness, dropping the model from 99.9% to 62.7% on ARC-AGI-3.
  • Three papers on September 9 extended the finding: Zhang et al. measured harness-induced variance at 7.8x model-induced variance; ByteDance's HarnessDev found Opus-generated harnesses dropped from 69.3 to 33.0 on SWE-Pro when Gemini executed them; and Amazon-Microsoft's SPACE raised ScienceWorld from 35.9% to 67.2% by training agents to batch safe sequential actions.
The gist

Published benchmark scores conflate model capability with harness design, making it difficult to compare models on standalone performance. The same model can appear near-perfect or modest depending on surrounding infrastructure.

08
Day 5

CoreWeave's first Vera Rubin NVL72 racks

  • On September 8, CoreWeave recapped the delivery of the first production Nvidia Vera Rubin NVL72 racks by Dell Technologies, which CoreWeave called 'a historic moment,' and said further hardware-related announcements are forthcoming.
  • Michael Dell had announced the milestone on social media, with CoreWeave's engineering and data center teams bringing up the units.
The gist

Michael Dell described the delivery as the world's first production Vera Rubin NVL72 shipment. CoreWeave has indicated more hardware announcements are forthcoming.

09
New

OpenAI ChatGPT Images 2.5 release

  • OpenAI released ChatGPT Images 2.5 on September 8, 2026, citing up to 50% lower latency than Images 2.0 and improvements to detail, lighting, texture, multi-turn instruction-following, and subject preservation.
  • The release ships two API variants: GPT-Image-2.5 Flare, the default for everyday generation, and GPT-Image-2.5 Sunburst, a slower, higher-control option for production creative work.
  • Adobe integrated the model into Firefly as a launch customer, the first named enterprise deployment.
  • Pricing, architecture details, and numerical benchmarks comparing the two variants remain unpublished.
The gist

The Adobe Firefly integration extends Images 2.5 into an established enterprise creative tool at launch, broadening reach beyond direct ChatGPT users. Speed improvements and two differentiated API models affect developers and creators choosing how to deploy image generation at scale.

10
New

Mistral AI's Samsung-led Series D

  • Mistral AI closed a €3 billion Series D on September 8 with Samsung Electronics as lead investor and Scaleup Europe Fund and PSG Equity as co-leads, lifting the valuation above €21 billion from €11.7 billion in September 2025.
  • Capital is earmarked for frontier research, training compute, and Mistral's first French data centre, with expansion planned across Europe, the Middle East, Asia, and North America.
  • Mistral also confirmed a July deal with Microsoft to provide AI computing capacity to Microsoft's clients.
The gist

The round nearly doubles Mistral's valuation from a year prior and funds a data centre build and international expansion that Mistral says will position it as a sovereign, cost-efficient AI alternative for organisations managing their own data and compliance constraints.

11
New

Google DeepMind's AlphaGenome Atlas

  • Google DeepMind launched AlphaGenome Atlas on September 8, a database of AI-predicted molecular effects for all 9 billion possible single-nucleotide variants in the human genome.
  • The 1-petabyte dataset, more than 30 times larger than the AlphaFold Database, introduces an AVI score combining AlphaGenome and AlphaMissense to rank variants by pathogenicity across coding and non-coding regions.
  • Collaborators used the AVI score to identify a previously overlooked DNM1 gene variant linked to epileptic encephalopathy, confirmed experimentally, and Atlas found 22% more non-coding genetic associations in UK Biobank data from over 54,000 participants.
The gist

The resource covers every possible single-nucleotide variant in the human genome, including non-coding regions that are difficult to assess for pathogenicity with existing tools. Early applications to UK Biobank data and rare disease cases produced validated findings, suggesting the dataset can extend what genome-wide analyses detect.

What is moving now · Every edition · Every story

The daily email

Want this in your inbox?

I send one email each morning with the stories that moved. If you would rather just read here, that works too.

Subscribe free