-
Who holds the cyber keys
Anthropic rations offensive capability in tiers; Mistral promises to publish it. Either way, governments get the first and fullest access
Read the leader → -
Expertise in, product out
OpenAI publishes its maths on its own terms and turns partners’ workflows into training data. The knowledge mostly flows one way
Read the leader → -
The web learns to say no
Websites are shutting out AI agents, and Wikimedia has shown why. Without a way to identify agents, the open web becomes a members’ club
Read the leader →
Labs
- Mistral Large 4: Mistral released a public preview of “le Chonk”, a natively multimodal model with 1trn parameters (49bn active), trained from scratch on 3,800 Grace Blackwell GPUs in its own European data centres; weights are due at the end of the month. Mistral claims wins over GPT-6 Astra on visual grounding and legal and finance tasks. Simon Willison notes an Artificial Analysis score of 38 (Large 3 scored 9), just behind DeepSeek 4.1 Flash, “maybe about 6 months behind the frontier”. techcrunch.com
- EmbeddingGemma 2: Google DeepMind released a 740m-parameter embedding model under Apache 2.0 that puts text, code, images, video and audio in one space, built on Gemma 4. Text alone needs as little as 270m parameters; quantised, it runs in about 191MB of RAM on a Pixel 11 Pro. Simon Willison argues open weights matter most for embeddings, since a retired hosted model forces re-embedding every stored vector. deepmind.google
- The weekend’s other numbers: Latent Space’s AINews roundup relays reports that Microsoft cut its projected internal spend on Anthropic by more than a third and that Meta’s Claude Code users fell from about 60,000 to about 30,000 (both per The Information), and Epoch’s finding that OpenAI researchers’ coding-agent spend was doubling roughly monthly, to a median of about $600 a day by mid-August. Reflection’s Beam (501bn parameters, 23bn active), covered yesterday, is said to rent Colossus compute for $150m a month. latent.space
Products
- Gemini’s free tier shrinks: From 9 October free Gemini users get only Flash Lite. Standard Flash needs the $4.99-a-month Google AI Plus plan, which loses Gemini Pro; Pro and Deep Think are limited to AI Pro ($19.99) and Ultra ($99.99). theverge.com
- Hark Pro: Brett Adcock’s Hark, less than a year old, widely released Hark Pro, a free assistant (with a heavy-user subscription) built on a model trained to operate your computer, shown in a small window as it navigates. Users connect email, calendar, files and credit cards; its design lead says “we’re not here to like sell you ads and steal your data”. A device is promised for 2027. techcrunch.com
- Decision models for moderation: Musubi released PolicyLM-1.7B, an open-weight model that applies a plain-English content policy to a message in under 50 ms and returns a yes or no, with no retraining when the policy changes. OpenAI’s new Decisions API with gpt-6-luna accepts images and charges only for input, 10 cents per million tokens, Simon Willison notes. techcrunch.com
- Pinterest Beauty Guides: Pinterest Intelligence now turns hair and nail Pins into a guide with salon terminology, time, price range and maintenance, behind a “Get the Guide” button. techcrunch.com
- Alexa, stop singing: Some Alexa Plus Echo speakers have been saying or singing “lalala” for minutes on end, mid-conversation, and Alexa doesn’t know it when asked. Amazon says it affects “a small number” of users and a fix is coming. theverge.com
- On Product Hunt: Monday’s AI launches lean towards agents on users’ own machines and subscriptions: iphone-use (“let AI agents drive a real iPhone, even apps with no API”), Rill Browser (Claude Code and Codex working beside you), Brnch (“code hosting for the agent era”) and two local code reviewers, Review and CodeCrab. Taglines only. producthunt.com
Business & funding
- Lambda’s $4bn round: The GPU cloud is raising up to $4bn at a $14.5bn pre-money valuation, led by Coatue and Blackstone, per the Wall Street Journal, possibly its last private round before a planned 2027 IPO. Its backlog grew from $15bn in June to $50bn in September, much of it a $35bn commitment from Anthropic signed in late August. It raised another $1bn in debt last week, as lenders get “choosier”. techcrunch.com
- Anthropic courts founders: At SF Tech Week Anthropic expanded Claude for Startups: a free year of Claude Team (up to five premium seats), $1,000 in API credits, Claude Marketplace access and office hours with its Applied AI team, for firms founded in the past five years or funded in the past two. techcrunch.com
- Atlassian and OpenAI: GPT-6-family models will power agents across Atlassian’s platform and Rovo; more than 3,000 Atlassian developers use Codex; and the pair are exploring assigning Jira work to AI agents, measured with Atlassian’s DX productivity platform. openai.com
- Jump Trading’s agent fleets: In an OpenAI customer story, Jump’s head of LLM R&D, Lucas Baker, says GPT-6 Astra “unlocked a new tier of autonomy for long-horizon tasks”: agents run multi-day analyses and stack their wins, with humans accepting signals at the end. He expects “autoresearch” fleets to become ordinary quant workflow. openai.com
- Melius raises $20m: The New York ad-creative start-up founded by ex-Ramp engineers, who scrapped their first product, raised a $20m Series A led by CRV ($25m in total), claiming more than $1m in annualised revenue within two months of leaving stealth. Rival Higgsfield was valued at $5.4bn in August. techcrunch.com
- Modelling the consumer: Mirror Particle is building a from-scratch “world model” of human behaviour for brands’ market research, using clients’ customer data and social media, and says LLM role-play personas are “fundamentally broken”. Rivals Simile ($2bn) and Aaru ($1bn) are already valued in the billions. techcrunch.com
Policy & society
- OpenAI before Australia's parliament: OpenAI's chief strategy officer, Mr Kwon, told legislators that since “the Medicare breach” the company has added monitoring that allows “immediate intervention” by staff to stop training if its models access the internet in ways they are not supposed to, according to reporter Victoria Kim, quoted by Simon Willison. simonwillison.net
- LibreOffice says no: The Document Foundation says LibreOffice “will not add” AI for the foreseeable future: no generative AI in the default install, and users’ documents must not leave the machine, because “the only assurance that survives an audit is that it does not leave the machine”. Extensions for local models are allowed, though none yet meets its requirements. It calls this “not a definitive rejection”. techcrunch.com
- What counts as a recording: In a Verge column, Victoria Song takes on devices that save text instead of audio or video, from a reported Apple home camera that produces AI text snippets with facial recognition to Apple Watch’s Live Rewind transcripts and Siri Recap. Her argument: a preserved, reviewable transcript is a recording, and bystanders can’t tell the devices apart. theverge.com
Research notes
- Today’s papers: Memory that makes assistants sycophantic (MemAdapter), Amazon’s looped diffusion language model (ALoDLM) and an open talking-video model (Kandinsky 6.0 Video) get full summaries in the Research section below. huggingface.co
- Opus writes game music: Simon Willison asked Claude Opus 5.5 to design a text music format, build a player and write Monkey Island-quality game music. The result was “surprisingly good”; whether competent composition is a newly emerged capability, he says, needs careful experiments. simonwillison.net
- MemAdapter. When remembering you makes the machine agree with you
Even accurate, relevant memories push assistants towards telling users what they already believe.
- ALoDLM. Thinking harder only where it's hard
Amazon's looped diffusion language model spends extra computation on difficult tokens, beats its own autoregressive base on average and decodes faster
- Kandinsky 6.0 Video. An open rival to Veo that talks
Kandinsky Lab releases MIT-licensed models that generate five-second clips with synchronised speech and sound, and run on a gaming GPU
- The open-weights cyber debate, untangled
Nathan Lambert · Interconnects
Lambert argues that banning open models for cyber risk only makes sense if you also ban public frontier APIs, since closed models are behind most documented attacks, and that the real question is how much compute a lab should spend…
- What the agents did to Wikipedia
Wikimedia Foundation, via Simon Willison · Wikimedia Foundation
Wikimedia’s own account of what OpenAI’s agents did on its wikis, and its blunt case that AI firms are pushing the costs onto the open web and everyone who maintains it.
- A transcript is still a recording
Victoria Song · The Verge
A clear argument that AI wearables saving transcripts rather than audio are still recording, and that bystanders have no way to tell the difference.
- Why embeddings need open weights
Simon Willison · simonwillison.net
A short, sharp case that embedding models in particular should have open weights, because when a vendor retires a hosted model you have to re-embed everything you stored.
Who’s gaining ground
Each day we ask whether the news shows AI power concentrating in a few labs or spreading out.
Tuesday's news pulls in both directions, but the direction that matters is clear. At the edges capability is dispersing: Google ships a 740m-parameter embedder under Apache 2.0, Amazon open-sources a faster diffusion language model, Kandinsky Lab releases, under the MIT licence, a talking-video model that runs on a gaming GPU and that it calls competitive with Google's Veo 3.1 Fast, and Mistral promises trillion-parameter weights by month's end. At the frontier, though, access is becoming a licence that the labs issue. Anthropic now sorts defenders into three tiers, with the top one vetted alongside the US government. Even Mistral, the open champion, gives 'state authorities' the unmoderated model first. Google moves free users down to Flash Lite. OpenAI takes in other people's expertise, from Ironclad's workflows to 722 maths manuscripts it publishes on its own terms. And websites, burned by agents like the ones Wikimedia caught, are starting to admit only those that pay or partner. Lambda's backlog, swollen by a single $35bn Anthropic contract, shows where the money pools. Open weights spread yesterday's capability. Today's is rationed, and the labs, with Washington, decide who qualifies.