Start with this article

Coding agents differ not by model but by the architectural bets of the harness. Reading the source of Codex, OpenCode, Pi and my own agent to tell load-bearing bets from ones you write to delete.
July 17, 2026
Strong articles that should not disappear in the archive

01 / 03
August 17, 2026

Part two of the prompt caching series. We dig into the vLLM and paged attention source to find the byte-level reason beh...
June 26, 2026

Prompt caching cuts input token costs by 10x and keeps agent unit economics afloat, yet it breaks without a single error...
June 10, 2026
A quick way into the part of the archive you need