Read the full issue →Download PDF
-
Defense as distribution
On the same day Anthropic rewrote the rules for deceptive agents and put Claude beside CrowdStrike inside power plants, it also made its strongest models free for open-source scans
Thursday was a busy day at Anthropic. The company published a 2026 Usage Policy refresh, effective 12th November, that consolidates bans on fake accounts and fabricated news into a single Do Not Engage in Deceptive Campaigns or Artificial Activity section, renames the elections rules Do Not Undermine Democratic Processes, drops the blanket ban on personalized campaign targeting (nonprofits writing ballot notices in other languages were caught; deceptive targeting remains banned elsewhere), and spells out stop-buttons for Claude when it is wired to hardware that can injure. Most of this, Anthropic says, restates enforcement already under way.
Read the leader → -
Afraid to ask outside
Three fired OpenAI safety researchers say ordinary outside collaboration is now punishable; the company insists on misconduct—either way, monitorability work now sits under fear
Jasmine Wang, Tomek Korbak and Mikita Balesni, the three safety researchers OpenAI fired last week, published an open letter on Thursday denying the company’s claim that they mishandled sensitive information and warning that colleagues are now “afraid to speak.” They say AI safety work depends on “close collaboration with outside experts,” and that terminations “executed and communicated so abruptly” chill the open culture OpenAI once prized. They deny involvement in a leak to The Information about less-monitorable architectures in newer models, and deny engaging outside parties beyond their job mandates.
Read the leader → -
The coworker with an inbox
Google’s Gemini agent gets its own email and audit trail for enterprises first—the personal-agent race is being won inside the workplace graph
At a Google Cloud event on Thursday, Gemini crossed from answering questions to “getting things done.” The new enterprise agent takes objectives, not just instructions: it plans work, loads skills, and connects to Workspace, Microsoft 365, Slack, Jira, Confluence, Git, BigQuery, Databricks, Postgres, Snowflake and any Model Context Protocol server inside or outside the network. By default it picks a model; users can override, including to Anthropic’s Claude, with open-source and private models promised later. A tasks inbox shows thinking, subagents and progress. Sundar Pichai noted more than a billion monthly Gemini users and that nearly 90% of Fortune 100 firms already use Gemini Enterprise.
Read the leader →
Labs
Products
- Gemini becomes a coworker techcrunch.com
- Google Foresight is a free, experimental Mac note-taker that transcribes meetings entirely on-device, a local-first shot… theverge.com
- Goodfire’s “inside-out” monitors read a model’s activations instead of re-reading its output with a second LLM techcrunch.com
- Natura’s $99 ring is a press-to-talk button for whichever personal agent you use, shipping December–January with… techcrunch.com
Business & funding
- OpenAI’s revenue, recounted techcrunch.com
- Arena, the model leaderboard, is worth $3.1bn after a $200m Series B led by Lightspeed and… techcrunch.com
- Manus raises more than $500m led by Boyu and IDG, its first round since Beijing forced… techcrunch.com
- Oracle goes big on OpenAI, with 130k ChatGPT and 95k+ Codex seats openai.com
- LegalOn halves its Codex bill openai.com
Policy & society
- Anthropic rewrites its rulebook anthropic.com
- OpenAI exposes its first Category-5 influence operation openai.com
- Fired safety researchers hit back techcrunch.com
- USA Today sues OpenAI for more than $250m, alleging it copied “hundreds of thousands” of articles theverge.com
- China will not slow down newsletter.semianalysis.com
- Long-WAM. Memory helps a robot only if it can imagine the future
NVIDIA’s Long-WAM shows longer visual history lifts success when video pretraining is causal—and runs in 107 ms on a consumer GPU
- Agent Oversight EBG. Not whether the agent finished—whether you can find what it changed
Evidence-Grounded Behavior Graphs improve localization of consequential decisions across eight models, and help Codex disclose silent changes
- Recursive Game Creator. Playable is not the same as fun
A Designer–Builder–Player–Reviewer harness scores player experience, lifting GameCraft to 77.89 and GameASG strict success to 53.2%