Archives
All the articles I've archived.
The Practices Everyone Says Help Coding Agents, Measured
Forced TDD, CLAUDE.md behavioral guidelines, token-saving skills: recent experiments measured them all, and most didn't survive contact with the data. Measure before you adopt a vibe.
Your First Reader Isn't Human
The web, docs, and CLIs are being redesigned for a new primary reader. Bot traffic passing humans, llms.txt, AXI, and why I write internal docs in English for AI.
A New Bar for AI Coding Benchmarks: The Question FrontierCode Asks
Cognition's new benchmark FrontierCode asks not 'does it pass the tests' but 'would a maintainer merge this PR.' Why the top model scores just 13.4% on the Diamond tier, and the gap between code that runs and code that merges.
The Things Summer Changed in Me
Summer isn't a season to endure. A reflection on how the sharpened senses and slowed-down time inside the heat carve our memories.
A Return After Seven Years: Why Jack Clark's Staged Release Is Back
Jack Clark, who designed the GPT-2 staged release in 2019, stands at the center of Project Glasswing in 2026. Tracing how the same logic returned with different evidence.
The Limits a Name Creates, and the Expansion Beyond It
When Claude Code does work beyond coding, how should its name change?
The Paradox of LLM Advancement
Updated:The easier writing becomes, the more we forget what we wrote. A look at cognitive debt in the AI era and the paradox of digital natives.
Winter Has Just Arrived
A small cold and a great cold are waiting for us. This winter, I will keep gathering beautiful sentences.
Grow Up to Be Someone's Sorrow
I am glad that you are my sorrow. So when you grow up, be sure to become someone's sorrow.
The Story of the Cat and the Boy
If you truly want to be loved purely, you had better keep a few cookie crumbs in your pocket.
Eternal Sunshine
Which is more beautiful, memory or love? Memory, of course. Memory lasts longer, and that is why it is more beautiful.