Latest signal September 20, 2026 · afternoon edition
Four of today's stories turn on the same gap, that deleting a thing and isolating a thing both take far more machinery than the button offering them implies.
01 / The wire
Recent briefings
-
September 20, 2026 · afternoon
Four of today's stories turn on the same gap, that deleting a thing and isolating a thing both take far more machinery than the button offering them implies.
-
September 20, 2026 · morning
Every performance number published about System One decision models this week came from a party selling the conclusion, and two of the primary artifacts say in their own text that the numbers do not transfer.
-
September 19, 2026 · afternoon
Four days after one lab shipped a model that only returns a decision, the ecosystem produced a browser clone, an open-weight rival with a priority claim, and a framework integration, and none of them has published a task-outcome comparison against the LLM it replaces.
-
September 19, 2026 · morning
Three separate groups attacked the token bill of long-horizon agents inside 48 hours, each at a different layer of the stack, and not one of them claims the cheaper output is more correct.
-
September 18, 2026 · afternoon
Five separate actors published evidence this week that the parts nobody picked on purpose, the context manager, the image decoder, the build dependency, the rate limiter, are where both the remaining performance and the entire blast radius now live.
-
September 18, 2026 · morning
The judgment calls buried inside coding harnesses, risk gating and compaction and model routing, are being unbundled into a cheap external decision model, and the community rebuilt them in seventy-two hours.
-
September 17, 2026 · afternoon
Agent memory and agent instructions are converging on plain files a human can read and Git can diff, and four vendors published an instance of that in 48 hours.
-
September 17, 2026 · morning
The layer between the model and the task, the harness and the handoff artifacts it writes, is where this week's cost and risk numbers landed, from a doubled bill for the same success rate to compaction summaries that carry instructions nobody wrote.
02 / Under the surface
Latest analysis
-
OpenAI's Measurement Pixel Leaks Your Visitors Before Its Own Privacy Code Runs
A third-party SDK cannot decline to send a cookie it has already sent, because the browser attaches credentials to the script request that…
-
Jev as a Judge: The Most Repeatable LLM in LangChain's Benchmark Was Also the Least Accurate
Repeatability and accuracy move in opposite directions across the three LLM judges in LangChain's benchmark, so low variance is not a proxy…
-
The Agent Cache That Hit 5 Percent, and the Release Note That Said So
A cache keyed on exact strings is worth nothing in a system whose input is speech, and the only reason anyone knows the number is that the…
-
CUA-S1 Ships a Model Card That Disqualifies Its Own Launch Benchmark
The reusable artifact in Cua's computer-use release is the model card's release checklist and abstention-aware metric set, not the…
-
What "Comparable Performance" Actually Means in an AI Efficiency Claim
Three headline efficiency results published in the same 48 hours use three different comparison shapes, and the word comparable in one of…
-
Supermemory Says MIT in the LICENSE File and 10,000 Documents in a Release Note
The enforceable limit on supermemory's self-hosted server lives in a release note and in the running binary rather than in the repository's…
-
Decision Models Report High Confidence on Inputs They Cannot Read
A decision model's confidence score measures how sharply the probability mass concentrates among the options you supplied, which makes it a…
-
Cloudflare's security-audit-skill Will Refuse to Start Rather Than Run a Thin Audit
The reusable machinery in Cloudflare's security-audit-skill is not its attack classes but its refusal machinery, a budget gate that…
04 / Coverage map
Topics we track
Claude Code 60 OpenAI 25 Anthropic 20 Codex 17 Agent Skills 15 Hugging Face 13 DeepSeek Harness 12 Model Context Protocol 12 Kimi K3 11 LangChain 8 MCP 8 METR 8 Claude Code auto mode 7 GitHub Copilot 7 GPT-5.6 Sol 7 MCP 2026-07-28 7 GPT-5.6-Cyber 6 Ollama 6 Anthropic Frontier Red Team 5 Claude Fable 5.1 5 Claude Opus 5 5 GLM-5.3 5 GPT-6 Astra 5 grok-build 5