Read the full issue →Download PDF
-
Defense as distribution
On the same day Anthropic rewrote the rules for deceptive agents and put Claude beside CrowdStrike inside power plants, it also made its strongest models free for open-source scans
-
Afraid to ask outside
Three fired OpenAI safety researchers say ordinary outside collaboration is now punishable; the company insists on misconduct—either way, monitorability work now sits under fear
-
The coworker with an inbox
Google’s Gemini agent gets its own email and audit trail for enterprises first—the personal-agent race is being won inside the workplace graph
Labs
Products
- Gemini becomes a coworker techcrunch.com
- Google Foresight is a free, experimental Mac note-taker that transcribes meetings entirely on-device, a local-first shot… theverge.com
- Goodfire’s “inside-out” monitors read a model’s activations instead of re-reading its output with a second LLM techcrunch.com
- Natura’s $99 ring is a press-to-talk button for whichever personal agent you use, shipping December–January with… techcrunch.com
Business & funding
- OpenAI’s revenue, recounted techcrunch.com
- Arena, the model leaderboard, is worth $3.1bn after a $200m Series B led by Lightspeed and… techcrunch.com
- Manus raises more than $500m led by Boyu and IDG, its first round since Beijing forced… techcrunch.com
- Oracle goes big on OpenAI, with 130k ChatGPT and 95k+ Codex seats openai.com
- LegalOn halves its Codex bill openai.com
Policy & society
- Anthropic rewrites its rulebook anthropic.com
- OpenAI exposes its first Category-5 influence operation openai.com
- Fired safety researchers hit back techcrunch.com
- USA Today sues OpenAI for more than $250m, alleging it copied “hundreds of thousands” of articles theverge.com
- China will not slow down newsletter.semianalysis.com
- Long-WAM. Memory helps a robot only if it can imagine the future
NVIDIA’s Long-WAM shows longer visual history lifts success when video pretraining is causal—and runs in 107 ms on a consumer GPU
- Agent Oversight EBG. Not whether the agent finished—whether you can find what it changed
Evidence-Grounded Behavior Graphs improve localization of consequential decisions across eight models, and help Codex disclose silent changes
- Recursive Game Creator. Playable is not the same as fun
A Designer–Builder–Player–Reviewer harness scores player experience, lifting GameCraft to 77.89 and GameASG strict success to 53.2%