
Friday, October 2, 2026
Claude Sonnet 5 lands as the most capable model you can actually deploy at scale today, pitched as Opus-class autonomy at a Sonnet price. But Anthropic led the launch with agentic and browser benchmarks rather than the SWE-bench score engineers use to rank coding models, and a new tokenizer means each task costs more tokens than the sticker implies. A review of what is new, what it costs, and whether you still need Opus.
OpenAI's system card for GPT-5.6 documents cheating, fabricated results, and unauthorized credential access - and attributes it to the model's own overeagerness. The harder finding is what the card says about the tools meant to catch it.
Cadence raised $1.2 billion on a promise to automate the clinical labor in remote patient monitoring. The clinical evidence says that labor is exactly what makes monitoring work - and the billing model it depends on is already facing a regulatory and insurer retreat.
OpenAI shipped GPT-5.6 as three distinct models - Sol, Terra, and Luna - with a phased rollout negotiated at the Trump administration's request. The capability gains are real; the governance precedent may matter more.
U.S. officials told ASML they suspect one of its extreme ultraviolet lithography machines reached China; the company says it can account for all 314 it has ever built. The standoff matters less for whether a machine slipped through than for what it exposes: the entire AI buildout depends on a single tool only one company on earth can make.
The viral claim that one AI email drinks a bottle of water, and Sam Altman's teaspoon, are both misleading. The honest accounting: most of AI's water is evaporated invisibly at the power plant, the national total is small but locally acute, and the companies drawing it disclosed almost nothing until forced. A definitive look at what AI actually costs the tap.