-
Defense as distribution
On the same day Anthropic rewrote the rules for deceptive agents and put Claude beside CrowdStrike inside power plants, it also made its strongest models free for open-source scans
Read the leader →Thursday was a busy day at Anthropic. The company published a 2026 Usage Policy refresh, effective 12th November, that consolidates bans on fake accounts and fabricated news into a single Do Not Engage in Deceptive Campaigns or Artificial Activity section, renames the elections rules Do Not Undermine Democratic Processes, drops the blanket ban on personalized campaign targeting (nonprofits writing ballot notices in other languages were caught; deceptive targeting remains banned elsewhere), and spells out stop-buttons for Claude when it is wired to hardware that can injure. Most of this, Anthropic says, restates enforcement already under way. The same day it launched the Anthropic Cyber Mission. The Critical Infrastructure Defense Program puts frontier Claude models, on-site engineers and threat research into the hands of the firms operators already trust for operational technology: Accenture, Booz Allen, CrowdStrike, Deloitte, Dragos, Hitachi, Insane Cyber, Nozomi Networks, Palo Alto Networks, PwC and Rockwell Automation. The pitch is that power grids, water plants and transport networks cannot simply be taken offline to patch, and that state…
-
Afraid to ask outside
Three fired OpenAI safety researchers say ordinary outside collaboration is now punishable; the company insists on misconduct—either way, monitorability work now sits under fear
Read the leader →Jasmine Wang, Tomek Korbak and Mikita Balesni, the three safety researchers OpenAI fired last week, published an open letter on Thursday denying the company’s claim that they mishandled sensitive information and warning that colleagues are now “afraid to speak.” They say AI safety work depends on “close collaboration with outside…
-
The coworker with an inbox
Google’s Gemini agent gets its own email and audit trail for enterprises first—the personal-agent race is being won inside the workplace graph
Read the leader →At a Google Cloud event on Thursday, Gemini crossed from answering questions to “getting things done.” The new enterprise agent takes objectives, not just instructions: it plans work, loads skills, and connects to Workspace, Microsoft 365, Slack, Jira, Confluence, Git, BigQuery, Databricks, Postgres, Snowflake and any Model Context Protocol server…
Labs
Anthropic’s Cyber Mission puts frontier Claude, on-site engineers and threat research beside CrowdStrike, Palo Alto Networks, Rockwell and others to defend grids, water and transport. A free, opt-in OSS Scanner sends unreviewed model-generated vulnerability reports to open-source projects. anthropic.com
Products
Google’s enterprise agent gets its own Workspace identity, email address and audit trail, and connects to Slack, Jira, M365, data warehouses and MCP servers. Businesses get it first. techcrunch.com
Business & funding
The Financial Times says OpenAI told investors annualised revenue is “approaching $50 billion”, about $20bn below a circulating $70bn figure that tried to match Anthropic’s partner-inclusive counting. IPO talk has slipped to early 2027. techcrunch.com
Policy & society
A 2026 Usage Policy, effective 12th November, folds bans on fake accounts and fabricated news into one section on deceptive campaigns and drops the blanket ban on personalised campaign targeting. Models wired to hardware that can injure must now have a stop-button and a safe state. anthropic.com
- Long-WAM. Memory helps a robot only if it can imagine the future
NVIDIA’s Long-WAM shows longer visual history lifts success when video pretraining is causal—and runs in 107 ms on a consumer GPU
- Agent Oversight EBG. Not whether the agent finished—whether you can find what it changed
Evidence-Grounded Behavior Graphs improve localization of consequential decisions across eight models, and help Codex disclose silent changes
- Recursive Game Creator. Playable is not the same as fun
A Designer–Builder–Player–Reviewer harness scores player experience, lifting GameCraft to 77.89 and GameASG strict success to 53.2%
- Beijing will not pace with you
SemiAnalysis
A census of 857 Chinese model releases: only 3.6% ever published a safety result.
- Probability without a novel
Timothy B. Lee · Understanding AI
The clearest plain-English guide to TypeSafe’s probability-only Jev model and why rivals are copying it.
- What the agents missed in the UV sky
Brice Ménard · Anthropic
A candid field note on Claude Science, including the artefacts the agents missed and a human caught.
- False fronts, Category 5
OpenAI
The richest primary-source case study yet of AI-assisted influence operations.
- When the drop is process, not a model
Zvi Mowshowitz · Don't Worry About the Vase
A dense weekly map of a week defined by math standards and safety firings; follow the receipts.
Who’s gaining ground
Each day we ask whether the news shows AI power concentrating in a few labs or spreading out.
Today’s announcements mostly consolidate. Anthropic routes Claude into critical infrastructure through incumbent contractors, Google gives its agent a badge inside the corporate graph, OpenAI shows how much it can see of ChatGPT’s covert users, and Arena becomes the paid scoreboard all of them share. The dispersal is at the edges: Manus refinanced in China, copies of Jev, and research harnesses that anyone can run. Distribution and oversight are concentrating; capability still leaks outward.