Beat: rag
13 pieces filed under rag, newest first.
-
ripwire Hands Coding Agents a Repo Map Instead of grep. Its Most Convincing Number Is the One That Got Worse.
Ripwire earns trust not with its 52x headline but by re-running its own head-to-head, publishing a corrected margin of 1.46x instead of the 1.75x its older tables…
-
OKF Agent Memory Puts Your Agent's Memory in Git. The Cost Is Buried in the Word BM25
Okf-agent-memory trades semantic recall for lexical recall and prices the trade as a latency win, so the reviewable-memory benefit is real but arrives with a retrieval…
-
Utopia's Append-Only Decision Ledger Runs as the Role That Can Delete It
Utopia's append-only decision ledger is enforced by Postgres triggers that its default single-role deployment is privileged enough to drop, so the audit guarantee is…
-
Perplexity Cites the Sites That Made 215,128 Machine-Written Buying Guides
An audit of 7,534 citations behind AI product recommendations found six in ten pointing outside the 100,000 most-visited sites, with three top-ten sources belonging to…
-
Briefing · September 3, 2026 · afternoon
Three agent launches in three days ship the same primitive, a human confirmation in front of the irreversible step, while a measurement of the retrieval layer those…
-
MemTrapBench Says Your Agent's Memory Is Making It Worse
Every memory framework MemTrapBench tested scored worse than the same model with memory switched off, which means the missing experiment in most agent stacks is not a…
-
OpenViking Turns Agent Memory Into a Directory You Can Walk
OpenViking's real contribution is not retrieval accuracy but retrieval evidence: a bad answer leaves a directory path you can read instead of a similarity score you…
-
firecrawl/anydoc: One Document Model Behind Fourteen Office Formats
Anydoc's real contribution is that every one of its fourteen formats parses into the same document model and renders through the same serializer, which is why a bug…
-
pdf-inspector: Firecrawl Says 54% of Your PDFs Never Needed OCR
Pdf-inspector's real argument is that roughly half the documents in a typical pipeline are already machine-readable and get sent to OCR anyway, and its own benchmark is…
-
book-to-skill Compiles a Technical Book Into an Agent Skill, Then Deletes the Book
Book-to-skill compiles a book into an instruction file your agent obeys, and the two rules that make it cheap and legally comfortable (never copy the author's words,…
-
VitaBench 2.0 Ran Three Agent Memory Architectures Against the Same Tasks, and Agentic Memory Won Half of Them
VitaBench 2.0's leaderboard shows agentic memory beating full context for 14 of 27 model entries and losing for every top scorer, which makes memory-architecture choice…
-
code-review-graph Argues Against Its Own Headline Number, and That's the Reason to Trust It
Code-review-graph's most copyable feature isn't the graph (code-graphs are a trending commodity) but that its README argues against its own headline number — it leads…
-
Briefing · July 20, 2026 · morning
The industry agreed agents should never touch source material directly, only curated projections (credential broker, knowledge compiler, code-graph layers), and…