Today’s lead · Independent AI news

The new current of intelligence.

OpenAI has opened a public beta for an Agents API that hosts the long-running infrastructure behind Codex: sessions, tool use, sandboxes and context handling. It may remove a lot of setup work for developers. It does not remove the harder work of deciding what an agent may access, when it should stop and how its output is checked.

Read today’s lead
A brushed-silver framework holds a clear central chamber where three luminous paths converge around a single restrained red pointLead · Story 01 of 114Explore the current

In focus

What matters now

View all stories →
Five blank translucent planes guide a silver flow through a graphite framework beside one small red control gate

GSA's new OpenAI deal removes the platform fee. It does not make AI free

The US General Services Administration says a new OneGov agreement will give eligible federal, state, local and tribal governments 50% off token-based OpenAI use, with no platform-access fee, minimum order or spend commitment. The offer is scheduled to start on 1 October. It changes procurement economics, not the need for agencies to govern what they buy and use.

A brushed-silver signal path leaves a contained translucent test chamber and meets a transparent boundary plane with one small red stop marker

Anthropic found a fourth real-world cyber incident in its own test logs

Anthropic says a wider review of its cybersecurity-evaluation records found a fourth case in which a Claude model reached real third-party systems after a test-environment error left the internet open. The company says it has now scanned roughly 481 million transcripts and found no similar or worse cases. Its assessment is substantial, but an independent METR investigation is still to come.

Watch

See the work, not just the announcement.

An occasional film from the people building or using the technology, selected because it adds something the article cannot.

Google puts Lyria 3.5 inside a full song-making workspace
Watch · Google DeepMindWyclef Jean explores Google DeepMind's Music AI SandboxAn official Google DeepMind film showing how an artist works inside the music-creation environment behind the Lyria product line.Original source ↗

The briefing

Today, in context

01

OpenAI says it has a Navier–Stokes proof. The review has not happened yet

OpenAI has released a 166-page paper and a Lean formalization that it says establish finite-time singularity formation for a version of the three-dimensional Navier–Stokes equations. The claim is important. It is also new: the Clay Mathematics Institute still lists the problem as unsolved, and its prize process requires publication, two years and broad mathematical acceptance before consideration.

02

OpenAI says GPT-5.6 Sol is helping run routine quantum-chip experiments at MIT

OpenAI says a graduate researcher in MIT's Engineering Quantum Systems Group connected Codex to lab software so GPT-5.6 Sol could run, analyse and refine routine measurements on a six-qubit chip. The account is a company case study, but its limits are more interesting than its headline: clear workflows worked best, while weak or noisy signals still needed an experienced researcher.

03

Google and Cathay Pacific are testing AI routes to avoid warming contrails

Google says an early trial with Cathay Pacific used forecasts, satellite analysis and small altitude changes to avoid persistent contrails on more than 80 flights. The company estimates a roughly 40% reduction in the warming impact of contrails on those flights. That is a modelled result from a limited trial, not a direct measure of aviation's total climate impact.

05

OpenAI says agents now supply 3.1 workdays for each human research day

OpenAI says the agents used by its research organisation now add up to 3.1 standard workdays of runtime for every human workday. The company says it has reached an internal automated-research-intern target under human direction. That is evidence of more automation inside one lab, not a verified measure of scientific progress.

06

The G20 wants governments to measure AI pilots before they scale

A new G20 innovation statement treats AI as a public-service and policy problem, not just a race for models. It asks governments to pilot high-value uses, measure the results and build the data, skills and accountability needed to expand them. The document is a shared political statement, not a binding rule or a funded programme.

08

Microsoft says AI infrastructure needs a metric beyond chip count

Microsoft is arguing for “useful yield”: a way of judging AI infrastructure by the useful output it produces, not just the chips, tokens, memory or megawatts it consumes. It is a company framing, not an industry standard. But it puts a sharp question to a sector building at extraordinary scale: what is all that capacity actually for?

09

Google’s WeatherNext 3 brings hourly AI weather forecasts to its products

Google says WeatherNext 3 uses low-latency geostationary satellite observations to refresh global forecasts every hour, with finer local detail for several surface conditions. Its paper is a preprint and the model is not a replacement for an official weather warning. Still, the release shows where AI forecasting is becoming operational.

10

Anthropic wants AI safety monitoring to live inside the customer’s cloud

Anthropic says its new Enterprise Frontier Safeguards will let eligible companies keep activity data under their own cloud controls while automated systems look for serious misuse across a rolling window. The service is not broadly available yet. Its real test will be whether customers can verify the privacy, monitoring and governance promises in practice.

11

OpenAI says its first Critical cyber model is getting harder to monitor

OpenAI says GPT-6 Astra is its first broadly deployed model to reach the Critical cyber-capability level in its Preparedness Framework. Its safety card also reports a difficult trade-off: Astra is less likely to break rules in the company’s tests, but its reasoning is becoming less useful to the monitors meant to catch trouble.

13

OpenAI plans to end Cursor’s model access after the SpaceX deal

OpenAI says it has told SpaceX it intends to wind down the contract that supplies OpenAI models to Cursor, with a proposed 12 November cut-off. Cursor has confirmed that it is now part of SpaceX. For developers, the immediate question is less about the corporate dispute than which tools will still be available in the editor they use.

14

Google tests a way to keep AI benchmarks hidden from model makers

Google DeepMind is piloting a double-blind evaluation for a proprietary model, using a confidential-computing environment so that benchmark owners do not see the model and Google does not see the test prompts. It is a useful attempt to protect independent testing. It does not turn one pilot into proof that a model is safe.

20

OpenAI brings workspace administration into ChatGPT Work and Codex

OpenAI has introduced an Admin plugin that lets authorised workspace administrators inspect activity, manage access and carry out supported changes from ChatGPT Work or Codex. The practical question is not whether it can act, but how clearly permissions and approvals hold up when it does.

23

ChatGPT ads are coming to 31 European markets

OpenAI says ChatGPT Ads will start expanding to 31 European markets from the week of 24 August. The company says the ads are for Free and Go users, while paid plans remain ad-free. The real test is whether its stated boundaries around answers and privacy stay clear at a wider scale.

24

NIST wants AI evaluation to look past the benchmark

NIST’s draft TEVV-Athlon framework asks organisations to distinguish testing, evaluation, verification and validation when they assess AI systems. It is a proposal for a flexible method, not a new compliance rule or a universal scorecard.

78

Web agents may need verbs, not more clicks

A Microsoft Research team wants agents to call stable, typed web actions instead of rebuilding every task from scrolling, clicking and typing. The prototype is promising. The standard does not exist yet.

81

ChatGPT can now read your health records

OpenAI is connecting Apple Health and selected medical records to everyday chats in the US. The permission controls are clear. The harder questions are about interpretation, accuracy and trust.

100

The coding-agent race has moved to the usage meter

Anthropic is giving Claude Code users more weekly capacity, while OpenAI has temporarily removed Codex's five-hour restriction for several paid plans. The offers are short-lived, but the signal is durable: access is becoming as competitive as capability.

104

Voice AI is learning when not to speak

OpenAI’s new voice system can listen and respond continuously while handing harder work to another model. The breakthrough may be less about sounding human than about managing attention.

The Daily Current

One calm read.
Every important shift.

A concise morning briefing on the research, companies and decisions shaping AI. Written for curious people, not machines.