Writing
Articles, architecture notes, project logs, and engineering reflections. Filter by type, search by keyword, or browse by tag.
Answering an operational question was costing a local model a 30-50K token dump. I built a compiled memory layer that does it in 2-5K, and blocks itself when it cannot cite a source.