Model Current

Policy

Rules, safety and the public decisions shaping artificial intelligence.

A small graphite control core directs three separate silver paths through individual red checkpoints across a dark field

Anthropic’s new threat report is a warning about workflow, not AI autonomy

Anthropic says it disrupted attempted misuse of Claude across seven harm areas between December 2025 and August 2026. Its account describes models being used inside tool-using attack workflows, while people still set targets and reviewed outcomes. The cases are substantial company evidence. They are not an independent measure of how often such misuse succeeds or a forecast of fully autonomous attacks.

A tall graphite gateway holds forty-five small silver spheres behind a single narrow red threshold against a black field

ChatGPT now has a four-month EU deadline. Designation is not a verdict

The European Commission has designated ChatGPT a Very Large Online Search Engine under the Digital Services Act after the service declared at least 45 million average monthly EU users. The designation triggers extra systemic-risk duties by January 2027. It does not mean the Commission has found that ChatGPT broke the law.

Five blank translucent planes guide a silver flow through a graphite framework beside one small red control gate

GSA's new OpenAI deal removes the platform fee. It does not make AI free

The US General Services Administration says a new OneGov agreement will give eligible federal, state, local and tribal governments 50% off token-based OpenAI use, with no platform-access fee, minimum order or spend commitment. The offer is scheduled to start on 1 October. It changes procurement economics, not the need for agencies to govern what they buy and use.

A ring of translucent graphite policy planes surrounds a small silver test chamber, with measured white paths and one restrained red threshold on a deep black field

The G20 wants governments to measure AI pilots before they scale

A new G20 innovation statement treats AI as a public-service and policy problem, not just a race for models. It asks governments to pilot high-value uses, measure the results and build the data, skills and accountability needed to expand them. The document is a shared political statement, not a binding rule or a funded programme.

A silver polyhedral core floats inside a black observation chamber, where a white signal and a narrow red sightline pass through layered smoked-glass planes

OpenAI says its first Critical cyber model is getting harder to monitor

OpenAI says GPT-6 Astra is its first broadly deployed model to reach the Critical cyber-capability level in its Preparedness Framework. Its safety card also reports a difficult trade-off: Astra is less likely to break rules in the company’s tests, but its reasoning is becoming less useful to the monitors meant to catch trouble.

Two silver forms send fine pale strands into opposite sides of a tall dark translucent chamber, where the paths remain separated by red-lit boundaries

Google tests a way to keep AI benchmarks hidden from model makers

Google DeepMind is piloting a double-blind evaluation for a proprietary model, using a confidential-computing environment so that benchmark owners do not see the model and Google does not see the test prompts. It is a useful attempt to protect independent testing. It does not turn one pilot into proof that a model is safe.

A silver path moves through four transparent planes above a dark graphite maze and ends at one red-lit mineral surface

NIST wants AI evaluation to look past the benchmark

NIST’s draft TEVV-Athlon framework asks organisations to distinguish testing, evaluation, verification and validation when they assess AI systems. It is a proposal for a flexible method, not a new compliance rule or a universal scorecard.