Skip to main content

Blog

Writing on the infrastructure layer of LLM systems — inference cost, serving, retrieval, evals, and observability. Mostly what I’ve measured, and what surprised me.

2026