
OpenAI has published the first performance results for Jalapeño, its custom inference chip. The company says it can improve throughput, latency and energy efficiency at once. The figures are detailed, but they are still OpenAI’s measurements ahead of deployment.

OpenAI has introduced an Admin plugin that lets authorised workspace administrators inspect activity, manage access and carry out supported changes from ChatGPT Work or Codex. The practical question is not whether it can act, but how clearly permissions and approvals hold up when it does.

A new Hugging Face analysis finds that attention around frontier open models and practical adoption are different things. Its data suggests smaller, older models still do much of the routine work, with one important limit: it describes activity on the Hub, not all AI use.

OpenAI says ChatGPT Ads will start expanding to 31 European markets from the week of 24 August. The company says the ads are for Free and Go users, while paid plans remain ad-free. The real test is whether its stated boundaries around answers and privacy stay clear at a wider scale.

In 1,902 controlled coding runs, the agent labelled coordinator did not become a communication hub or reliably improve success. The shape of the task mattered more.

Across six open-model update pairs and six benchmarks, no single inference-time signal reliably found the regressions hidden inside a higher average score.

The new workhorse targets coding and agents. Its launch price is half the permanent rate, and Google is already putting it inside a 24/7 personal agent.

In 9,840 simulated supply-chain negotiations, most agents found efficient agreements. The weaker models were far more likely to accept terms that broke their own profit rule.

Luna now costs $0.20 per million input tokens, while a new Fast mode charges twice as much to run Sol sooner. OpenAI's own documentation has not fully caught up.

Dozens of companies have joined a push for shared AI security infrastructure. Its first concrete release is NOOA, a framework whose own warning explains why open tooling is only part of the answer.

The EU may provide up to €10 billion and hopes to draw more than €20 billion from private investors. The tender is real; much of the money and the energy are not yet in place.

Project Perception enters public preview on 3 August, starting with software vulnerability management. Microsoft has shared strong benchmark and cost numbers. Real-world performance and the exact limits on automated action are still open.

OpenAI found that 43.5% of occupation-specific messages in its sample matched tasks linked to another occupation. The pattern is revealing. It does not tell us whether the work was good, used or reviewed.

Presence is not a do-it-yourself chatbot kit. OpenAI is selling a managed deployment with scoped access, testing, approvals and human handoffs for voice and chat workflows.

Gemini 3.6 Flash is the workhorse, Flash-Lite is built for cheap high-volume tasks, and a cyber specialist will stay limited to governments and trusted partners.

Both companies made the case on the same day. One starts with work completed, the other with compute used. Neither has produced a shared standard.

The AI platform says an autonomous agent ran a multi-stage intrusion through a malicious dataset. Public models appear untouched, but the company is still checking whether customer or partner data was affected.

A new AI Office report says Europe’s research base is not the main weakness. The urgent gaps are compute, energy and growth capital. The report is expert advice, not Commission policy.

Anthropic is giving Claude Code users more weekly capacity, while OpenAI has temporarily removed Codex's five-hour restriction for several paid plans. The offers are short-lived, but the signal is durable: access is becoming as competitive as capability.

Apple alleges that OpenAI benefited from confidential hardware information brought over by former staff. The claims have not been tested in court, but the case exposes a growing pressure point: people can change companies; trade secrets cannot.

The model is built for tasks that mix language, images and tool use. The interesting test now is whether developers can turn that breadth into dependable products.

Demand climbed sharply in 2025, and the build-out is accelerating. The constraint is no longer only chips: it is grids, permits, cooling and time.