GET THE AUTOPSY ➔

OpenAI and Anthropic are selling the same next step for AI agents: more of them. Claude Code now forks subagents by default, and Sol Ultra fans a problem across up to 64. Google Research ran the controlled test, and the answer is a split: more agents help work that breaks into independent pieces and hurt work that runs as one dependent chain, by up to 70%. Which one your task is decides whether the swarm is an upgrade or a tax.

The pitch is that a swarm of parallel subagents beats one worker. The most rigorous study of it, 260 configurations across six benchmarks from Google Research and academic collaborators, says it depends entirely on task shape. On parallelizable work, coordination helps a lot, up to +81%. On sequential, dependent work like planning, every multi-agent setup they tested got worse, by 39 to 70%. A lot of day-to-day coding, debugging a dependent chain, planning a change, is the second kind, and defaulting the swarm on only helps if it fans out for independent subtasks rather than splitting one line of reasoning.

01THE CLAIM
"Parallel subagents are the next capability step for AI agents: the two biggest coding agents now lean on them, with Claude Code forking subagents by default and OpenAI's Sol Ultra fanning a problem across up to 64 concurrent subagents, on the premise that more agents produce better results." [SOURCE ↗]
TRUE, BUT6 SOURCES · LIVE 2026-08-25
OPENAI TRACK RECORD33 CLAIMS · 39/100 BS RATE →
-70%WORST MULTI-AGENT DEGRADATION ON A STRICTLY SEQUENTIAL TASK (PLANCRAFT)
39-70%RANGE BY WHICH EVERY MULTI-AGENT VARIANT DEGRADED SEQUENTIAL REASONING
+81%WHERE MORE AGENTS DO HELP: A PARALLELIZABLE TASK (FINANCE-AGENT)
64CONCURRENT SUBAGENTS OPENAI'S SOL ULTRA CAN FAN A PROBLEM ACROSS; 4 IS THE REAL DEFAULT
OpenAI and Anthropic are selling the same next step for AI agents: more of them. Claude Code now forks subagents by default, and Sol Ultra fans a problem across up to 64. Google Research ran the controlled test, and the answer is a split: more agents help work that breaks into independent pieces and hurt work that runs as one dependent chain, by up to 70%. Which one your task is decides whether the swarm is an upgrade or a tax.
02THE CHECK

THE CLAIM. parallel subagents are the next leap for AI agents. On August 13 Claude Code turned subagent forking on by default, OpenAI's Sol Ultra fans a task across up to 64 concurrent subagents, and the industry heuristic driving all of it is, in the words of the researchers who tested it, the belief that adding specialized agents will consistently improve results.

THE CHECK. Google Research and academic collaborators ran the first controlled study large enough to answer it, 260 configurations across six agentic benchmarks and five architectures, holding tools, prompts, and compute fixed. The verdict is not that multi-agent is fake. It is that task structure decides. On a parallelizable benchmark like Finance-Agent, coordination delivered up to +81%. On a strictly sequential one like PlanCraft, every multi-agent variant they tested degraded performance by 39 to 70%. Their own summary: coordination improves performance on parallelizable tasks but degrades it on sequential ones. The open question for the two coding agents pushing subagents hardest is which pattern their default triggers: fanning out to gather context in parallel is the winning case, but splitting a dependent job like debugging or planning across coordinating agents is the losing one, and much of what developers hand these tools is the dependent, tool-heavy kind.

03SAY THIS IN THE MEETING · 📸 SCREENSHOT IT
"Both big coding agents now push parallel subagents as the win, but the controlled Google Research study of 260 configs found more agents help parallelizable tasks by up to 81% and hurt sequential ones like planning by up to 70%. Task structure decides, not agent count, so a default that fans out for independent subtasks can help while one that splits a dependent chain is a tax."

The agent story of the season is parallelism. On August 13, Claude Code shipped a release whose changelog states plainly that 'Subagent forking is now on by default,' turning a swarm of parallel Claude instances from an expert opt-in into the standard behavior. OpenAI is pushing the same idea from t

🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNT

You just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 6 sources with quotes and screenshots, and our on-record call.

This story is a stable, citable object. If you can falsify a verdict, tell us. Corrections are loud here.