Start with two full samples. Sign in for this edition’s free story. Sources and corrections stay open.
What the pen marks mean
rememberThe takeaway: what to remember
figureThe number that matters
evidenceThe evidence line
rulingThe ruling
claimThe claim that fails
Each mark is drawn once, as you reach it. Numbers are only circled when they come from the story’s own figures.
01 / ModelsTrue, but
Google says Gemini 4 Argon beat OpenAI and Anthropic across benchmarks. The independent index has it tied with GPT-6 Astra and five points behind Claude Opus 5.5.
Google, Gemini 4 Argon launch post ·
Google says Gemini 4 Argon scored significantly higher than GPT-6 Astra and Anthropic's Fable and Opus across a variety of benchmarks, with a new top score on DeepSWE v1.1 (77.9%).
The reality check
Google picked the tests; some are run by third parties such as Vals. On Artificial Analysis's independent index, Argon scores 53, matching GPT-6 Astra, and The Decoder reports Claude Opus 5.5 at 58.
Why care? If you are choosing a model, Argon is a credible peer of GPT-6 Astra, not a clear leader, and the price edge rests on introductory rates. Test it on your own tasks before switching.
Take this with youGoogle says Gemini 4 Argon beat OpenAI and Anthropic on benchmarks, but the independent Artificial Analysis index has it tied with GPT-6 Astra at 53 and behind Claude Opus 5.5 at 58.
Open the evidence4 source pages +
The claim we checked
Google's Gemini 4 Argon scored significantly higher than OpenAI's GPT-6 Astra and Anthropic's Fable and Opus models across a variety of AI benchmarks.
These are quoted receipts, not a count of independent investigations. Several reports may rely on the same original source.
Google claims that Argon scored significantly higher than OpenAI’s GPT-6 Astra and Anthropic’s Fable and Opus models across a variety of AI benchmarks.
OpenAI says dots are always-on agents "built to handle everything". On its own launch page, the always-on part uses read-only tools that cannot send a message.
OpenAI, Introducing dots (DevDay 2026) ·
OpenAI launched dots at DevDay as "remarkably capable, always-on agents built to handle everything", running on GPT-6 Astra for Pro, Business Premium and Enterprise users.
The reality check
The same page says background "proactive research" uses tools restricted to read-only, which cannot send messages or change app content; acting follows approval rules; specialist dots are focused enterprise pilots; and users should review consequential work. TechCrunch says much of it was already possible through Codex.
Why care? Dots ship as a careful, permissioned assistant, not an agent that handles everything. The safety in it is the read-only default and the approval rules you set, so set them.
Take this with youOpenAI pitches dots as always-on agents, but its own launch page says background work uses read-only tools that cannot send messages.
Open the evidence5 source pages +
The claim we checked
OpenAI says its new dots are "remarkably capable, always-on agents built to handle everything", working toward users' goals around the clock.
These are quoted receipts, not a count of independent investigations. Several reports may rely on the same original source.
The FTC was already looking into the companies prior to OpenAI's disclosure that its high-powered AI models hacked into AI company Hugging Face in July.
The FTC plans to issue formal demands for information and compel testimony from executives at top AI developers, including Anthropic, OpenAI and the research group METR
AI 'needs $6 trillion a year' to justify the data-centre boom. The number is a $1.5 trillion spending forecast divided by an assumed ratio of about a quarter.
Bain & Company, 7th Global Technology Report ·
Bain & Company's 2026 Global Technology Report says funding AI's compute demand would require $6 trillion in annual revenue by 2031, with existing consumer and enterprise AI at $1.2 trillion to $1.8 trillion and $4.2 trillion left to new categories.
Members · 30 days free
See what the evidence actually shows.
Members read the full check on every story: what the evidence shows, why it matters to you and the one line to take with you. Every past edition, re-verified, and the Receipts Pack come with it. A$89 a year, about A$0.24 a day.
First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge.
Open the evidence4 source pages +
The claim we checked
AI needs $6 trillion in annual revenue by 2031 to justify the global data-centre boom, per Bain & Company's 2026 Global Technology Report.
These are quoted receipts, not a count of independent investigations. Several reports may rely on the same original source.
AI infrastructure is being built well ahead of the demand curve and funding it sustainably will require adding approximately 1% to the annual global GDP growth rate