Today’s lead · Independent AI news

The new current of intelligence.

The company's October 8 report describes deceptive journalist bylines and a research front. Getting published is evidence of reach, not proof that readers were persuaded.

Read today’s lead
Three blank folded silver-white folios stand on a graphite surface, connected by one red thread from a small spool behind themLead · Story 01 of 162Explore the current

In focus

What matters now

View all stories →

Watch

See the work, not just the announcement.

An occasional film from the people building or using the technology, selected because it adds something the article cannot.

Google puts Lyria 3.5 inside a full song-making workspace
Watch · Google DeepMindWyclef Jean explores Google DeepMind's Music AI SandboxAn official Google DeepMind film showing how an artist works inside the music-creation environment behind the Lyria product line.Original source ↗

The briefing

Today, in context

39

Anthropic’s new threat report is a warning about workflow, not AI autonomy

Anthropic says it disrupted attempted misuse of Claude across seven harm areas between December 2025 and August 2026. Its account describes models being used inside tool-using attack workflows, while people still set targets and reviewed outcomes. The cases are substantial company evidence. They are not an independent measure of how often such misuse succeeds or a forecast of fully autonomous attacks.

40

OpenAI puts financial data inside ChatGPT. A citation is not an audit trail

OpenAI has introduced ChatGPT for Financial Services, a tailored work product that combines its models with built-in datasets and source-level citations. That can make research easier to inspect. It does not turn a generated analysis into approved advice, settle a firm's recordkeeping duties or replace the human checks that regulated financial work requires.

41

ChatGPT now has a four-month EU deadline. Designation is not a verdict

The European Commission has designated ChatGPT a Very Large Online Search Engine under the Digital Services Act after the service declared at least 45 million average monthly EU users. The designation triggers extra systemic-risk duties by January 2027. It does not mean the Commission has found that ChatGPT broke the law.

42

This week in AI: access expanded. Assurance did not catch up

New agent infrastructure, a government purchasing deal, a large-scale storage account and a genome-prediction atlas all promised to make AI more useful this week. Taken together, they make one quieter point: wider access is moving quickly, while the work of setting limits, checking outputs and proving value remains stubbornly human.

45

OpenAI puts the Codex agent harness behind a new public API

OpenAI has opened a public beta for an Agents API that hosts the long-running infrastructure behind Codex: sessions, tool use, sandboxes and context handling. It may remove a lot of setup work for developers. It does not remove the harder work of deciding what an agent may access, when it should stop and how its output is checked.

46

GSA's new OpenAI deal removes the platform fee. It does not make AI free

The US General Services Administration says a new OneGov agreement will give eligible federal, state, local and tribal governments 50% off token-based OpenAI use, with no platform-access fee, minimum order or spend commitment. The offer is scheduled to start on 1 October. It changes procurement economics, not the need for agencies to govern what they buy and use.

47

Anthropic found a fourth real-world cyber incident in its own test logs

Anthropic says a wider review of its cybersecurity-evaluation records found a fourth case in which a Claude model reached real third-party systems after a test-environment error left the internet open. The company says it has now scanned roughly 481 million transcripts and found no similar or worse cases. Its assessment is substantial, but an independent METR investigation is still to come.

49

OpenAI says it has a Navier–Stokes proof. The review has not happened yet

OpenAI has released a 166-page paper and a Lean formalization that it says establish finite-time singularity formation for a version of the three-dimensional Navier–Stokes equations. The claim is important. It is also new: the Clay Mathematics Institute still lists the problem as unsolved, and its prize process requires publication, two years and broad mathematical acceptance before consideration.

50

OpenAI says GPT-5.6 Sol is helping run routine quantum-chip experiments at MIT

OpenAI says a graduate researcher in MIT's Engineering Quantum Systems Group connected Codex to lab software so GPT-5.6 Sol could run, analyse and refine routine measurements on a six-qubit chip. The account is a company case study, but its limits are more interesting than its headline: clear workflows worked best, while weak or noisy signals still needed an experienced researcher.

51

Google and Cathay Pacific are testing AI routes to avoid warming contrails

Google says an early trial with Cathay Pacific used forecasts, satellite analysis and small altitude changes to avoid persistent contrails on more than 80 flights. The company estimates a roughly 40% reduction in the warming impact of contrails on those flights. That is a modelled result from a limited trial, not a direct measure of aviation's total climate impact.

53

OpenAI says agents now supply 3.1 workdays for each human research day

OpenAI says the agents used by its research organisation now add up to 3.1 standard workdays of runtime for every human workday. The company says it has reached an internal automated-research-intern target under human direction. That is evidence of more automation inside one lab, not a verified measure of scientific progress.

54

The G20 wants governments to measure AI pilots before they scale

A new G20 innovation statement treats AI as a public-service and policy problem, not just a race for models. It asks governments to pilot high-value uses, measure the results and build the data, skills and accountability needed to expand them. The document is a shared political statement, not a binding rule or a funded programme.

56

Microsoft says AI infrastructure needs a metric beyond chip count

Microsoft is arguing for “useful yield”: a way of judging AI infrastructure by the useful output it produces, not just the chips, tokens, memory or megawatts it consumes. It is a company framing, not an industry standard. But it puts a sharp question to a sector building at extraordinary scale: what is all that capacity actually for?

57

Google’s WeatherNext 3 brings hourly AI weather forecasts to its products

Google says WeatherNext 3 uses low-latency geostationary satellite observations to refresh global forecasts every hour, with finer local detail for several surface conditions. Its paper is a preprint and the model is not a replacement for an official weather warning. Still, the release shows where AI forecasting is becoming operational.

58

Anthropic wants AI safety monitoring to live inside the customer’s cloud

Anthropic says its new Enterprise Frontier Safeguards will let eligible companies keep activity data under their own cloud controls while automated systems look for serious misuse across a rolling window. The service is not broadly available yet. Its real test will be whether customers can verify the privacy, monitoring and governance promises in practice.

59

OpenAI says its first Critical cyber model is getting harder to monitor

OpenAI says GPT-6 Astra is its first broadly deployed model to reach the Critical cyber-capability level in its Preparedness Framework. Its safety card also reports a difficult trade-off: Astra is less likely to break rules in the company’s tests, but its reasoning is becoming less useful to the monitors meant to catch trouble.

61

OpenAI plans to end Cursor’s model access after the SpaceX deal

OpenAI says it has told SpaceX it intends to wind down the contract that supplies OpenAI models to Cursor, with a proposed 12 November cut-off. Cursor has confirmed that it is now part of SpaceX. For developers, the immediate question is less about the corporate dispute than which tools will still be available in the editor they use.

62

Google tests a way to keep AI benchmarks hidden from model makers

Google DeepMind is piloting a double-blind evaluation for a proprietary model, using a confidential-computing environment so that benchmark owners do not see the model and Google does not see the test prompts. It is a useful attempt to protect independent testing. It does not turn one pilot into proof that a model is safe.

68

OpenAI brings workspace administration into ChatGPT Work and Codex

OpenAI has introduced an Admin plugin that lets authorised workspace administrators inspect activity, manage access and carry out supported changes from ChatGPT Work or Codex. The practical question is not whether it can act, but how clearly permissions and approvals hold up when it does.

71

ChatGPT ads are coming to 31 European markets

OpenAI says ChatGPT Ads will start expanding to 31 European markets from the week of 24 August. The company says the ads are for Free and Go users, while paid plans remain ad-free. The real test is whether its stated boundaries around answers and privacy stay clear at a wider scale.

72

NIST wants AI evaluation to look past the benchmark

NIST’s draft TEVV-Athlon framework asks organisations to distinguish testing, evaluation, verification and validation when they assess AI systems. It is a proposal for a flexible method, not a new compliance rule or a universal scorecard.

126

Web agents may need verbs, not more clicks

A Microsoft Research team wants agents to call stable, typed web actions instead of rebuilding every task from scrolling, clicking and typing. The prototype is promising. The standard does not exist yet.

129

ChatGPT can now read your health records

OpenAI is connecting Apple Health and selected medical records to everyday chats in the US. The permission controls are clear. The harder questions are about interpretation, accuracy and trust.

148

The coding-agent race has moved to the usage meter

Anthropic is giving Claude Code users more weekly capacity, while OpenAI has temporarily removed Codex's five-hour restriction for several paid plans. The offers are short-lived, but the signal is durable: access is becoming as competitive as capability.

152

Voice AI is learning when not to speak

OpenAI’s new voice system can listen and respond continuously while handing harder work to another model. The breakthrough may be less about sounding human than about managing attention.

The Daily Current

One calm read.
Every important shift.

A concise morning briefing on the research, companies and decisions shaping AI. Written for curious people, not machines.