My Local LLM Scored 6/6. It Was Wrong Every Time.
Six months of trying to make a 1.2B model useful, and the measurement mistakes I made along the way.
Read articleLocal models. Real tools. Honest measurements.
I’m Mark Hall, an AI engineer at GoEngineer. I write about local language models, agentic systems, evaluation, and the engineering between a convincing demo and dependable software.
Latest writing
Six months of trying to make a 1.2B model useful, and the measurement mistakes I made along the way.
Read articleOpen source project
A local-first AI assistant for Windows with permissioned tools, durable memory, visible evidence, and a hard stop button.