Anthropic Resignation and a 10% Risk Estimate — AI Digest

An Anthropic researcher resigned publicly, saying the labs are moving too fast toward superintelligence; the risk estimate for humanity is above 10%. GPT-6 Astra compiled a game on macOS at 120 FPS, and an open 2-billion-parameter model topped the under-4B ranking.
Today's highlights
An Anthropic researcher's resignation became the loudest AI-risk statement yet
Jacob Coxon, an Anthropic employee and former OpenAI researcher, resigned publicly — posting on social media that leading labs are moving too fast toward self-improving superintelligence and toward agents capable of running cyberattacks. By Ben's Bites' account, his post became probably the most-viewed AI safety communication ever. The reaction split not along "agree or disagree" but along "what to do about it": Yoshua Bengio wrote that warnings from frontier-lab researchers deserve to be taken seriously, David Shor called for government-mandated independent oversight, and several researchers publicly vouched for Coxon's credibility. Separately that same day, an Anthropic researcher put the chance of AI killing all humans within the next decade at more than 10%. Ben's Bites weekly roundup, Bengio's position, the call for oversight, the probability estimate and the debate.
GPT-6 Astra compiled and ran a game on macOS: 120 FPS after six hours
Developer Theo reported that GPT-6 Astra independently compiled and ran Super Smash Bros. Melee on macOS at 120 frames per second, working roughly six hours in a single loop. In parallel, Vals reported that Astra nearly saturated their unreleased computer-use evaluation by building a Minecraft Nether portal in under three hours with no task-specific harness. This is the week's second signal of the same kind: the model operates not on text about a program but on the program itself — mouse, windows, compiler. The practical meaning for work: tasks that used to require writing a script per interface now yield to a plain description of the goal, at the price of hours of machine time. The Melee demo, the Vals measurement.
A 2-billion-parameter model tops the open-weight ranking under 4B
OpenBMB open-weighted MiniCPM5-2B and claims 15 points on the Artificial Analysis Intelligence Index v4.2 — the best result among open models up to 4 billion parameters. Weights are on Hugging Face, code on GitHub. The practical interest of the discussion moved from benchmarks to pipelines: a model this size fits a "speech recognition → model → speech synthesis" chain on weak hardware that used to require a server. A related result landed alongside it: the author of the TAK quantization method reports that his version of Qwen3.8-27B scores 82.81% against 83.59% for the original at full precision, while occupying 15% of its size. Quantization is compression that coarsens the numbers inside a model: it gets several times smaller and faster, but usually loses quality. The MiniCPM5-2B release, the quantization write-up.
Apple A20 Pro: 115 GB/s of memory bandwidth in a phone — laptop M4 territory
Apple's A20 Pro chip gets a 7-core GPU, a doubled 32-core Neural Engine and roughly 115 GB/s of memory bandwidth — about 50% more than the A19 Pro and nearly the 120 GB/s of the laptop M4. Memory bandwidth is the speed at which the processor reads a model's weights; for local AI it matters more than clock speed, because the whole model is read for every answer. Per Notebookcheck's breakdown, the chip moves to TSMC's 2-nanometre process and gets a 96-bit LPDDR5X bus. The constraint is unchanged and it is not speed: the phone is still expected to ship with 12 GB of RAM, and the size of a local model is bounded by capacity, not bandwidth. The specifications breakdown.
LangChain shipped secret management for agents
LangChain Managed Deep Agents 0.7 introduced Connections — a mechanism that separates the agent's own access from the user's: some connections belong to the agent itself, others run through a person's OAuth sign-in. It answers a concrete problem: an agent handed a key to mail or spreadsheets keeps it in its own environment, and revoking that access separately from the agent is impossible. A VS Code update landed alongside it, with automation for recurring work and GitHub flows inside the agents window. Both teams' theme for the week is the same — not "smarter model" but "more predictable permissions". The Connections announcement, the VS Code update.
Numbers and facts
- 120 frames per second after ~6 hours of work — Astra compiled and ran Super Smash Bros. Melee on macOS (demo).
- Under 3 hours — building a Minecraft Nether portal on a computer-use evaluation, with no specialized harness (Vals).
- 15 points on the Artificial Analysis Intelligence Index v4.2 — MiniCPM5-2B, the best result among open models up to 4B (release).
- 82.81% against 83.59% at 15% of the size — TAK quantization of Qwen3.8-27B versus the original model at full precision (write-up).
- ~115 GB/s against the M4's 120 GB/s — memory bandwidth of the Apple A20 Pro and of the laptop chip (breakdown).
- $468 million and up to 10x HBM capacity — Kepler Compute emerged from seven years of stealth with its own fab and memory that does not depend on EUV (claim).
- 364 of 10,416 skills in the AI SKILLS catalog for Claude Code as of 11 September 2026 describe computer or browser control; agents are mentioned in the descriptions of 2,991 skills.
- More than 10% — the probability of AI killing all humans within the next decade, as estimated by an Anthropic researcher (the debate).
Different perspectives: are warnings from inside the labs a signal or a campaign?
The dispute is not about the facts but about how to read a public resignation. Nobody contests what happened: an Anthropic employee and former OpenAI researcher left publicly, saying the labs are moving too fast toward self-improving superintelligence.
For "this is a signal". Yoshua Bengio wrote plainly that warnings from frontier-lab researchers deserve to be taken seriously (Bengio). Several researchers publicly vouched for Coxon's credibility, among them Ethan Perez and Will Depue (Perez, Depue).
Neutral. David Shor moves the conversation from trust to structure: what is needed is mandatory independent government oversight, not faith in companies' good intentions (Shor). Framed that way, one person's motives stop being decisive.
Against. Parker Thayer described the episode as a politicized campaign bordering on a "psyop" (Thayer); Ben's Bites records that part of the audience is asking outright whether this is a conspiracy (roundup).
Tools and techniques
- Separate the agent's access from your own. Connections in LangChain Managed Deep Agents 0.7 let some connections sit on the agent's side and others run through the user's OAuth sign-in, so revoking access does not require taking the agent apart (announcement).
- Photon 2.2 extended local inference to nearly the whole NVIDIA fleet — A10/A10G, A100, 3090, L4, H100, B200 and RTX PRO 6000 Blackwell — plus a megakernel compiler upgrade: a unified kernel feeds the GPU more evenly when the CPU is busy (release). A curated set of open tools for local inference lives in the AI SKILLS Open Source section.
- An open format for AI-drawn diagrams. Ben's Bites highlights Eraser Diagrams — an open format in which a model draws and lays out diagrams instead of handing back a picture (issue). Ready-made agent scenarios live in the AI SKILLS automation templates catalog.
In brief
- OpenAI rolled back a failure that consumed banked resets in ChatGPT Work and Codex; affected users are promised replacement resets and apology emails (statement).
- OpenAI's training-data opt-out does not stack: the in-app setting and the privacy portal are two routes to the same switch, not two separate ones (clarification).
- Arena introduced GameDevBench — a benchmark of deterministic game-development tasks derived from real tutorials (Arena).
- StereoPolicy claims 3D perception for robot manipulation straight from a stereo pair, with no depth maps or LiDAR, beating RGB, RGB-D and PointNet on tabletop tasks (Lambda).
- Anthropic built an interactive model of the 2030 economy, and Codex gained a Unity plugin plus a collection of plugins for small businesses (roundup).
- "Artificial", the film about Sam Altman's firing and return, arrives at Christmas with Andrew Garfield playing Altman (roundup).
- iOS 28 is expected to ship a feature that proves a photo's authenticity (roundup).
- A biologist with a computational-biology doctorate says he synthesized PAC-3310, a compound designed by ChatGPT, in his garage; the work has not been independently verified (discussion).