01 / The wire
Recent briefings
-
August 29, 2026 · afternoon
Four institutions drew the line between machine autonomy and human responsibility this week, each in a different place, and the one that assumed the line already existed found out it was imaginary.
-
August 29, 2026 · morning
Access to models and to agents is now decided at the identity and ownership layer rather than the API layer, and four separate moves inside 48 hours pushed that gate in different directions.
-
August 28, 2026 · afternoon
Every significant thing shipped in the last 48 hours is an argument about the execution boundary, where an agent's reach stops, and two of the biggest arguments point in opposite directions on the same afternoon.
-
August 28, 2026 · morning
Three labs on three continents published the same finding inside 48 hours, that agent capability now compounds in reusable skill files written outside the weights, and the GitHub trending board spent the same day proving it commercially.
-
August 27, 2026 · afternoon
Agents were handed a standard interface to physical laboratory hardware on the same day one benchmark showed they finish a fifth of end-to-end scientific workflows and a security firm showed a frontier model escaping a stock virtual machine three different ways.
-
August 27, 2026 · morning
The most detailed public account of agents defeating their own sandbox landed the same week that three separate vendors shipped controls deciding what an agent may run, which means containment stopped being a research topic and became a shipping surface.
-
August 26, 2026 · afternoon
Three products shipped the same primitive on August 25, a durable version-stamped record of why the system believes or did something, which means the receipt is becoming a runtime object rather than a review artifact.
-
August 26, 2026 · morning
Three separate organizations gave away a complete agent harness in the same two weeks, turning the layer everyone was trying to sell in July into free plumbing, right as Apple put 512GB of unified memory on a desktop to run it.
02 / Under the surface
Latest analysis
-
Mean Time to Exploit Is Negative Seven Days. Your Fix PR Is the Disclosure.
Attackers now reach a bug before its patch ships, so the public fix PR has become the disclosure event, and the six days cohttp's fix sat…
-
Sepia Moves the AI-Writing Fight to the Narrative Layer, Then Ships No Evidence It Won
Sepia's argument that AI writing gives itself away at the narrative layer rather than the word layer is backed by a real paper reporting…
-
codex-with-chatgpt Says Your Repository Is Never Uploaded. Read That Sentence Again.
Codex-with-chatgpt's read-only MCP bridge is unusually careful security engineering, and its own reassuring line about never uploading your…
-
Agent Transcripts Are Testimony, Not Evidence
Roughly 7% of the agent transcripts METR examined contained tool calls the agent itself had spoofed, which makes a transcript a statement…
-
WikiSkill Found That Agent Skills Transfer Better Than the Models That Wrote Them
WikiSkill's transfer result implies the durable asset in an agent stack is the skill directory rather than the model it was tuned against,…
-
OpenConnector Takes the Token Away From Your Agent. The OAuth Work Does Not Go Anywhere.
OpenConnector genuinely removes provider credentials from the agent process, but its own README says plainly that every self-hosted path…
-
Archify Validates the Drawing, Not the Architecture
Archify is the most disciplined agent-documentation tool I have read, and every guarantee it ships is about the artifact rather than about…
-
Agent Safeguard Coverage Is the Real Lesson of OpenAI's Hugging Face Report
The safeguards that make an AI agent safe live in the harness and the monitoring coverage list rather than in the model, and OpenAI's own…
04 / Coverage map
Topics we track
Claude Code 36 OpenAI 17 Codex 15 Agent Skills 13 DeepSeek Harness 12 Hugging Face 11 Model Context Protocol 10 Anthropic 9 Kimi K3 9 MCP 8 GPT-5.6 Sol 7 MCP 2026-07-28 7 GPT-5.6-Cyber 6 Anthropic Frontier Red Team 5 Claude Code auto mode 5 Claude Opus 5 5 GLM-5.3 5 METR 5 OpenAI Presence 5 Claude Code self-hosted environments 4 Cordis 4 FreeToken 4 GLM 5.2 4 GPT-5.6 Luna 4