Writing

Mostly long-form notes on how LLM systems actually work underneath the API — serving engines, agent harnesses, retrieval and the parts of evaluation nobody enjoys. I write these to check my own understanding, so they tend to go one level deeper than they need to.