Skip to content
[bmdpat]

Writing

What fits, what it costs, and why it is slow

Local LLMs, GPU sizing, GGUF quants, and what AI agents cost. The local-AI numbers come from dated runs on hardware I own. Every other figure names its source.

Start here

The questions people ask most, answered.

Full archive (190)
Artifact preview for My 8B Model Failed a 400-Word Taskreal test output
6 min read

My 8B Model Failed a 400-Word Task

Three Llama 3.1 8B runs missed a 400-word floor. Here is the verifier-driven route that moved long-form synthesis to Gemma 4 26B.

Read the post

The AI agent build notes

Real costs, real tools, no fluff. One evidence-backed note on Friday when there is something worth sharing.

Get the requested artifact now, then at most one evidence-backed Local AI Lab Note on Friday when there is something worth sharing. One-click unsubscribe. No sponsored placements. Privacy.