Archive

Search

162 stories

A small graphite control core directs three separate silver paths through individual red checkpoints across a dark field

Anthropic’s new threat report is a warning about workflow, not AI autonomy

Anthropic says it disrupted attempted misuse of Claude across seven harm areas between December 2025 and August 2026. Its account describes models being used inside tool-using attack workflows, while people still set targets and reviewed outcomes. The cases are substantial company evidence. They are not an independent measure of how often such misuse succeeds or a forecast of fully autonomous attacks.

A stack of blank silver reference planes aligns beneath a graphite lens, with one small red verification mark on a dark field

OpenAI puts financial data inside ChatGPT. A citation is not an audit trail

OpenAI has introduced ChatGPT for Financial Services, a tailored work product that combines its models with built-in datasets and source-level citations. That can make research easier to inspect. It does not turn a generated analysis into approved advice, settle a firm's recordkeeping duties or replace the human checks that regulated financial work requires.

A tall graphite gateway holds forty-five small silver spheres behind a single narrow red threshold against a black field

ChatGPT now has a four-month EU deadline. Designation is not a verdict

The European Commission has designated ChatGPT a Very Large Online Search Engine under the Digital Services Act after the service declared at least 45 million average monthly EU users. The designation triggers extra systemic-risk duties by January 2027. It does not mean the Commission has found that ChatGPT broke the law.

Four silver pathways widen from a shared graphite core before each passes through a small separate red gate on a black field

This week in AI: access expanded. Assurance did not catch up

New agent infrastructure, a government purchasing deal, a large-scale storage account and a genome-prediction atlas all promised to make AI more useful this week. Taken together, they make one quieter point: wider access is moving quickly, while the work of setting limits, checking outputs and proving value remains stubbornly human.

A brushed-silver framework holds a clear central chamber where three luminous paths converge around a single restrained red point

OpenAI puts the Codex agent harness behind a new public API

OpenAI has opened a public beta for an Agents API that hosts the long-running infrastructure behind Codex: sessions, tool use, sandboxes and context handling. It may remove a lot of setup work for developers. It does not remove the harder work of deciding what an agent may access, when it should stop and how its output is checked.

Five blank translucent planes guide a silver flow through a graphite framework beside one small red control gate

GSA's new OpenAI deal removes the platform fee. It does not make AI free

The US General Services Administration says a new OneGov agreement will give eligible federal, state, local and tribal governments 50% off token-based OpenAI use, with no platform-access fee, minimum order or spend commitment. The offer is scheduled to start on 1 October. It changes procurement economics, not the need for agencies to govern what they buy and use.

A brushed-silver signal path leaves a contained translucent test chamber and meets a transparent boundary plane with one small red stop marker

Anthropic found a fourth real-world cyber incident in its own test logs

Anthropic says a wider review of its cybersecurity-evaluation records found a fourth case in which a Claude model reached real third-party systems after a test-environment error left the internet open. The company says it has now scanned roughly 481 million transcripts and found no similar or worse cases. Its assessment is substantial, but an independent METR investigation is still to come.

A silver fluid vortex narrows through a transparent graphite verification frame beneath a separate ring of light with one restrained red point

OpenAI says it has a Navier–Stokes proof. The review has not happened yet

OpenAI has released a 166-page paper and a Lean formalization that it says establish finite-time singularity formation for a version of the three-dimensional Navier–Stokes equations. The claim is important. It is also new: the Clay Mathematics Institute still lists the problem as unsolved, and its prize process requires publication, two years and broad mathematical acceptance before consideration.

A compact silver quantum-chip form sits inside six concentric translucent measurement paths that converge on a small red control gate

OpenAI says GPT-5.6 Sol is helping run routine quantum-chip experiments at MIT

OpenAI says a graduate researcher in MIT's Engineering Quantum Systems Group connected Codex to lab software so GPT-5.6 Sol could run, analyse and refine routine measurements on a six-qubit chip. The account is a company case study, but its limits are more interesting than its headline: clear workflows worked best, while weak or noisy signals still needed an experienced researcher.

A brushed silver route makes three precise altitude shifts through translucent graphite atmospheric layers, with a single restrained red threshold on a deep black field

Google and Cathay Pacific are testing AI routes to avoid warming contrails

Google says an early trial with Cathay Pacific used forecasts, satellite analysis and small altitude changes to avoid persistent contrails on more than 80 flights. The company estimates a roughly 40% reduction in the warming impact of contrails on those flights. That is a modelled result from a limited trial, not a direct measure of aviation's total climate impact.

Three fine white paths circle a single graphite research table and converge into one restrained silver aperture, with a small red human control point on a deep black field

OpenAI says agents now supply 3.1 workdays for each human research day

OpenAI says the agents used by its research organisation now add up to 3.1 standard workdays of runtime for every human workday. The company says it has reached an internal automated-research-intern target under human direction. That is evidence of more automation inside one lab, not a verified measure of scientific progress.

A ring of translucent graphite policy planes surrounds a small silver test chamber, with measured white paths and one restrained red threshold on a deep black field

The G20 wants governments to measure AI pilots before they scale

A new G20 innovation statement treats AI as a public-service and policy problem, not just a race for models. It asks governments to pilot high-value uses, measure the results and build the data, skills and accountability needed to expand them. The document is a shared political statement, not a binding rule or a funded programme.

A vertical sculpture of silver, graphite and translucent mineral layers concentrates a diffuse white stream into a single bright point, with one small red adjustment mark

Microsoft says AI infrastructure needs a metric beyond chip count

Microsoft is arguing for “useful yield”: a way of judging AI infrastructure by the useful output it produces, not just the chips, tokens, memory or megawatts it consumes. It is a company framing, not an industry standard. But it puts a sharp question to a sector building at extraordinary scale: what is all that capacity actually for?

A silvery three-dimensional contour field unfolds through a black space, with a fine white weather front moving across it and one restrained red calibration line

Google’s WeatherNext 3 brings hourly AI weather forecasts to its products

Google says WeatherNext 3 uses low-latency geostationary satellite observations to refresh global forecasts every hour, with finer local detail for several surface conditions. Its paper is a preprint and the model is not a replacement for an official weather warning. Still, the release shows where AI forecasting is becoming operational.

A dark graphite archive chamber holds suspended translucent data panes inside a silver outer frame, with one small red boundary line and no visible text

Anthropic wants AI safety monitoring to live inside the customer’s cloud

Anthropic says its new Enterprise Frontier Safeguards will let eligible companies keep activity data under their own cloud controls while automated systems look for serious misuse across a rolling window. The service is not broadly available yet. Its real test will be whether customers can verify the privacy, monitoring and governance promises in practice.

A silver polyhedral core floats inside a black observation chamber, where a white signal and a narrow red sightline pass through layered smoked-glass planes

OpenAI says its first Critical cyber model is getting harder to monitor

OpenAI says GPT-6 Astra is its first broadly deployed model to reach the Critical cyber-capability level in its Preparedness Framework. Its safety card also reports a difficult trade-off: Astra is less likely to break rules in the company’s tests, but its reasoning is becoming less useful to the monitors meant to catch trouble.

A brushed-silver track splits at a graphite mechanism into a short straight route and a longer glowing route through transparent planes, with a small red gate at the far end

Google gives Gemini 3.8 Flash a choice: answer fast or keep working

Google says Gemini 3.8 Flash keeps the introductory price of its predecessor while doing more multi-step reasoning and tool use on difficult tasks. Its new Cyber variant is more tightly restricted, which is the more important part of the release for anyone thinking about capable AI agents.

Two wide brushed-silver routes stop just short of each other across a narrow red-lit gap in a black spatial field

OpenAI plans to end Cursor’s model access after the SpaceX deal

OpenAI says it has told SpaceX it intends to wind down the contract that supplies OpenAI models to Cursor, with a proposed 12 November cut-off. Cursor has confirmed that it is now part of SpaceX. For developers, the immediate question is less about the corporate dispute than which tools will still be available in the editor they use.

Two silver forms send fine pale strands into opposite sides of a tall dark translucent chamber, where the paths remain separated by red-lit boundaries

Google tests a way to keep AI benchmarks hidden from model makers

Google DeepMind is piloting a double-blind evaluation for a proprietary model, using a confidential-computing environment so that benchmark owners do not see the model and Google does not see the test prompts. It is a useful attempt to protect independent testing. It does not turn one pilot into proof that a model is safe.

A transparent graphite control plane directs one incoming silver flow into three separate paths, with a small red checkpoint at its base

OpenAI brings workspace administration into ChatGPT Work and Codex

OpenAI has introduced an Admin plugin that lets authorised workspace administrators inspect activity, manage access and carry out supported changes from ChatGPT Work or Codex. The practical question is not whether it can act, but how clearly permissions and approvals hold up when it does.

A narrow red-lit threshold opens into a black field of suspended translucent panels linked by fine silver lines

ChatGPT ads are coming to 31 European markets

OpenAI says ChatGPT Ads will start expanding to 31 European markets from the week of 24 August. The company says the ads are for Free and Go users, while paid plans remain ad-free. The real test is whether its stated boundaries around answers and privacy stay clear at a wider scale.

A silver path moves through four transparent planes above a dark graphite maze and ends at one red-lit mineral surface

NIST wants AI evaluation to look past the benchmark

NIST’s draft TEVV-Athlon framework asks organisations to distinguish testing, evaluation, verification and validation when they assess AI systems. It is a proposal for a flexible method, not a new compliance rule or a universal scorecard.

A dark diagnostic instrument gathers many uneven silver symptom fragments into a layered white clinical map while a small red marker remains outside the final boundary

Google tested a symptom-checking AI with nearly 14,000 people

SymptomAI asked follow-up questions and produced useful lists of possible diagnoses in a large US study. The result is notable. So are the limits: clinicians judged transcripts gathered by the AI, and the main reference diagnoses were reported by participants.

Many loose silver contact points enter a dark instrument and emerge as three stable geometric action modules linked to a plain red verification disc

Web agents may need verbs, not more clicks

A Microsoft Research team wants agents to call stable, typed web actions instead of rebuilding every task from scrolling, clicking and typing. The prototype is promising. The standard does not exist yet.

Translucent health-data fragments pass through a dark metal permission aperture controlled by one small red latch

ChatGPT can now read your health records

OpenAI is connecting Apple Health and selected medical records to everyday chats in the US. The permission controls are clear. The harder questions are about interpretation, accuracy and trust.

Two silver capacity reels feed sweeping luminous work streams past a small red glass latch

The coding-agent race has moved to the usage meter

Anthropic is giving Claude Code users more weekly capacity, while OpenAI has temporarily removed Codex's five-hour restriction for several paid plans. The offers are short-lived, but the signal is durable: access is becoming as competitive as capability.

Two silver arcs exchange fine luminous currents around a suspended red pulse in a black field

Voice AI is learning when not to speak

OpenAI’s new voice system can listen and respond continuously while handing harder work to another model. The breakthrough may be less about sounding human than about managing attention.