On becoming a judge in the age of AI — and why that might be, in some domains, the same thing as becoming a creator.
Claude Code + Tmux + Modal ran nine hours of autonomous ML research. The results are real.
Twelve thousand games of The Resistance: Avalon say the mission board is the least informative surface in the game. What finds a liar is the talk—and Evil's most disciplined habit turns out to be a mistake it pays ten points for.
Read →A practical map of an agent setup — canonical config, skills worth building first, subagent delegation with git worktrees, a second model, a few plugins worth adding, a notes vault, and a feedback loop that writes corrections back into the skills — concrete enough to copy.
Read →GORGO is a cross-region LLM request router that scores cache locality, replica load, and wide-area network latency in one cost function—and learns the weights from the deployment’s own latency stream.
9 minRead →Why HDR so often looks wrong — and what the file is really telling your screen. Relative vs. absolute light, the PQ transfer function, and the Rec.2020 gamut, with drag-to-compare HDR demos.
13 minRead →Mechanistic interpretability tools that reveal how Claude's internal circuits execute reasoning, plan outputs, and implement safety mechanisms.
Read ↗On comprehending a mathematical proof as an ecstatic experience — a transfiguration from ambivalence and skepticism to conviction.
22 minRead →