How Penwright stitches whisper.cpp and llama.cpp together
A short technical tour of the on-device AI pipeline.
DraftingEssays on local AI, Mac tooling, and the math of typing vs. speaking. New posts land when we have something useful to say.
Privacy, latency, and the engineering bet on whisper.cpp + llama.cpp + Metal. The opinionated essay version.
An honest tour of our three-tier on-device cleanup pipeline, rules engine, small local model, EuroLLM 1.7B (native EN + DE), and why we don't reach for a 7B+ model for every sentence.
A short technical tour of the on-device AI pipeline.
DraftingWhat happens when typing disappears from the equation.
Drafting