All postsTag
Local Inference
2026
An Agent Reads My Email
Every morning at five, a local model reads yesterday's mail and writes me a briefing. How it works, what 217 days of numbers say, and what happens when spam is written for the model instead of for me.
Notes from the AI Gatherings, Vol. 1
First in a recurring series - what we compared notes on this week at the Friday morning AI gathering: cleaning data 700 calls at a time, the two economies of tokens, and building the tool before the feature.
The Machine I Can't Turn Off
A four-year-old Mac Studio runs a night shift of local agents - launchd as the runtime, open-weight models as the brains, git as the memory. What I've learned building it, and the economics thesis I'm testing.
Lab
Can a Local Model Read My Inbox? A Bake-off
I gave three on-device models the same job - turn each email into a structured record - and judged them against Claude Opus. Three models, four runs, because Qwen ran once with thinking on and once with it off. With thinking off, the smallest-feeling model won on faithfulness, classification and speed.
Getting good output from Lens
If your Lens recaps come out cluttered with reasoning preambles or hallucinated detail, the fix is almost always upstream of Lens. A field guide to picking a model and configuring LM Studio for clean daily snapshots.
Why personal AI belongs on hardware you already own
Apple just put two hardware engineers at the top of the company. The cloud labs are losing money on their best customers. From inside a hybrid practice, here's what the local share looks like and why it's growing.
A general-purpose MoE multimodal beat every dedicated vision model on my father's handwriting
I assumed a specialized vision model would win. I was wrong. A head-to-head on a hard handwriting corpus ended with the general-purpose MoE on top.
Running real models locally on a Mac Studio that isn't new anymore
How I run a multimodal LLM on four-year-old hardware to read a family archive without sending anything to the cloud.
What family archives are for, now that the AI can read them
I have three collections of family letters spanning a century. Until recently, reading them properly would have taken years. Now it takes an afternoon.