01 / The wire
Recent briefings
-
October 8, 2026 · afternoon
On October 8 Anthropic and LangChain each put the last check outside the model, a stop a person holds and an approval the model cannot write, while OpenAI's October 7 math withdrawals showed what happens when the check comes after publication.
-
October 8, 2026 · morning
On October 6 and 7 GitHub said agent-era volume had outrun its Git storage and its secret scanning, while Cloudflare and Anthropic showed the same fix pattern, moving trust out of plain-language instructions and into code that checks.
-
October 7, 2026 · afternoon
On October 7 Anthropic priced its new small model about 90 percent below Haiku 4.5 for prompts under 100K tokens and OpenAI wrapped GPT-6 in a UI compiler, while Nvidia's olympiad results and Armin Ronacher's Codemode essay both show the system around the model now decides the result.
-
October 7, 2026 · morning
On October 6 OpenAI disclosed about three hours of Pro-level thinking per math result while dropping output charges on its Decisions API, and Google shipped an embedding model sized for a phone, a widening gap between costly reasoning and cheap judgment that agent builders should design around.
-
October 6, 2026 · afternoon
On October 6 three labs set the terms of model capability three different ways, Anthropic by verifying who the user is, OpenAI by training a frontier model on one partner's workflows, and Mistral by marketing cyber tasks that closed models refuse, so what a model will do for you now depends on the terms around it as much as on the model.
-
October 6, 2026 · morning
Within two days an open-weight model arrived as an announcement without weights, agent-proposed semiconductors arrived with their key properties computed but unmeasured, and ChatGPT signed cartoons with the names of people who never drew them, so AI output is running ahead of the evidence that says who made it and whether it holds.
-
October 5, 2026 · afternoon
On October 5 OpenAI announced a statistical watermark for its own text while the Wikimedia Foundation described, with heavy hedging, traffic it attributes to suspected OpenAI agents, so provenance for AI words is shipping faster than provenance for AI actions.
-
October 5, 2026 · morning
Over the past week the vendor AI app became the contested surface, as OpenAI announced ad slots for ChatGPT image generation and a Florida arrest report, picked up on October 4, showed what Anthropic's safety review can send to police, while builders shipped tools that strip Apple's models out of macOS, lift ChatGPT's computer-use runtime into other harnesses, and point agents at other people's binaries.
02 / Under the surface
Latest analysis
-
tsc-rs Passes 181,711 Tests and Nobody Has Read the Code. The Test Suite Is Now the Spec
When an LLM ports a codebase that no human reads, the inherited test suite becomes the only spec, so tsc-rs's own Known problems list, not…
-
Docker Agent Lets You Pull an AI Agent Like an Image. Its Docs Say Not to Trust It Like One
Docker Agent lets anyone run an AI agent straight from an OCI registry, but its own docs call permissions client-side and not a security…
-
Claude Code Prompt Hooks Let Through What They Were Told to Block
Claude Code 2.1.294 fixed prompt and agent hooks written as instructions that allowed what they should block, which shows that a guard a…
-
Agent Payments With Stripe Link: The Token Approves the Money, Your Code Approves the Cart
A Link shared payment token is one-time and amount-bound but knows nothing about the cart, so an agent purchase is only as safe as the tool…
-
OpenAI's Math Repo Holds 722 Manuscripts, and Its Own Catalogue Formalizes a Fraction of Them
Openai/math's own formalization catalogue links 162 of its 722 manuscripts to a formalized main result and marks its review status…
-
OpenAI's Decisions API Bills Only What It Reads, and That Changes How You Write Agent Guards
OpenAI's Decisions API charges only for input tokens, at twice gpt-6-luna's regular input rate, so a decision guard's cost and accuracy…
-
anthropics/knowledge-work-plugins Is a 123-Plugin Marketplace Now
Anthropics/knowledge-work-plugins is now a 123-entry marketplace whose README still describes 11 plugins, and its own sales plugin declares…
-
Claude Haiku 5.5 Is a Migration, Not a Model Swap
Claude Haiku 5.5 lists at a tenth of Haiku 4.5's price for prompts under 100K tokens, but a straight model-ID swap breaks on sampling…
04 / Coverage map
Topics we track
Claude Code 80 OpenAI 33 Anthropic 25 Codex 18 Agent Skills 15 DeepSeek Harness 14 Hugging Face 13 Model Context Protocol 13 Kimi K3 11 MCP 11 GitHub Copilot 9 LangChain 9 METR 8 Claude Code auto mode 7 GPT-5.6 Sol 7 MCP 2026-07-28 7 GPT-5.6-Cyber 6 Ollama 6 Anthropic Frontier Red Team 5 Claude Fable 5.1 5 Claude Opus 5 5 GLM-5.3 5 GPT-6 Astra 5 grok-build 5