Latest signal September 17, 2026 · afternoon edition
Agent memory and agent instructions are converging on plain files a human can read and Git can diff, and four vendors published an instance of that in 48 hours.
01 / The wire
Recent briefings
-
September 17, 2026 · afternoon
Agent memory and agent instructions are converging on plain files a human can read and Git can diff, and four vendors published an instance of that in 48 hours.
-
September 17, 2026 · morning
The layer between the model and the task, the harness and the handoff artifacts it writes, is where this week's cost and risk numbers landed, from a doubled bill for the same success rate to compaction summaries that carry instructions nobody wrote.
-
September 16, 2026 · afternoon
The day's launches stopped asking whether an agent can do the task and started asking whether it will do it again, so the new products are measurements of repeatability and deterministic rails around the model rather than smarter models.
-
September 16, 2026 · morning
Tuesday's launches from Cloudflare, Anthropic, and TypeSafe all replace an all-or-nothing switch with a typed, scoped control, while the day's security story shows the old switches still leaking.
-
September 15, 2026 · afternoon
The same week the labs escalated the story that agents are becoming dangerous threat actors, three independent outside reads pushed back, and the gap between the catastrophe framing and the agents you can actually observe got wide enough to see through.
-
September 15, 2026 · morning
Supervision of agents is leaving the prompt and becoming a separate runtime component with its own veto, whether that is a per-command network allowlist, a guard model trained on execution events, or a validating agent that is never the one that found the bug.
-
September 14, 2026 · afternoon
Four desktop and OS vendors shipped in five days, and every one of them kept the harness and made the model the swappable part.
-
September 14, 2026 · morning
The product shipping this weekend was the workspace around the model, its files, memory, tools and provenance, and not the model itself.
02 / Under the surface
Latest analysis
-
Skills Over MCP vs Specialist Agents: Microsoft's Own Trace Says Fewer Calls, More Tokens
A specialist agent is usually a SKILL.md and three tools wearing a model, and Microsoft's own trace shows that removing the model halves…
-
PRAXIST Hit #1 on GitHub Trending Under a License Called Fair Source That Never Converts to Open Source
PRAXIST's license borrows the Fair Source name, drops the delayed open-source publication that fair.io's definition requires, and adds a…
-
HarnessTax Says Your Coding Agent Harness Costs 2x for 2 Points, and the Bill Starts on the First Call
Harness choice is a cost decision that has been sold as a feature decision, and HarnessTax shows the cheapest place to measure it is the…
-
funes Keeps the Original Passage: Why Hugging Face's Coding-Agent Memory Refuses to Summarize
Funes bets that coding-agent memory should be a rebuildable index of verbatim transcript passages rather than model-written summaries, and…
-
TypeSafe Jev Is a Frontier Model That Cannot Write a Sentence, and That Is the Point
Most decisions inside an agent loop are enums, and a model that returns typed probabilities instead of text turns the hallucinated tool…
-
An AI Agent Found a Live Admin Token in a 2023 Docker Image in 25 Minutes. Yours Is Probably Still There.
This week's scoped-token launches govern credentials minted from now on, but the credential that gets you was baked into a container's…
-
Open Code Review: Alibaba Put the Parts That Must Not Fail in Code, Not in the Prompt
Open Code Review's value is the deterministic layer that fences the LLM into writing comments, and its self-run benchmark, lower recall by…
-
The Consistency Gap: Your Agent's 77% Is Hiding a 24-Point Reliability Problem
An agent's average pass rate hides a consistency gap caused by near-tied token decisions flipping under platform noise, so the metric to…
04 / Coverage map
Topics we track
Claude Code 57 OpenAI 24 Anthropic 20 Codex 17 Agent Skills 15 Hugging Face 13 DeepSeek Harness 12 Model Context Protocol 12 Kimi K3 11 MCP 8 METR 8 Claude Code auto mode 7 GitHub Copilot 7 GPT-5.6 Sol 7 MCP 2026-07-28 7 GPT-5.6-Cyber 6 Ollama 6 Anthropic Frontier Red Team 5 Claude Fable 5.1 5 Claude Opus 5 5 GLM-5.3 5 GPT-6 Astra 5 grok-build 5 OpenAI Presence 5