The School Attached to the Lab Is Open Now
The model bake-off got its final standings. The training platform behind it is now free to read — every module, both games, no account required.
Writing
Notes on AI-assisted development, agentic workflows, and building software that ships. No thought leadership — see why.
The model bake-off got its final standings. The training platform behind it is now free to read — every module, both games, no account required.
Four fixed assignments, every current model, every effort setting, three attempts each — and the grades barely move while the bill moves by a factor of four hundred.
Anthropic's own telemetry shows people approve 93% of permission prompts — that's not caution, it's a reflex, and the fix is a policy you set before you're tired.
Context rot isn't a cliff a model falls off — it's a gradient, and the fix is curating what's in the window, not waiting for a collapse.
Chatbot, human-in-the-loop, agentic workflow, and full autonomy aren't stages you graduate through — they're positions you choose per task, and the vendors selling autonomy keep telling you not to default to it.
Model, harness, and tools are three different jobs — and the coding agent's terminal access is why it can do things a chatbot never could.
Research, Plan, Implement, Verify is just the ordinary shape of software work, finally said out loud — and the spec/plan mix-up inside it is costing teams more than they realize.
An agent can change forty files in ninety seconds and describe it in one confident sentence. That sentence is a claim, not a record — the diff is the only account of the work that nobody wrote.
Hunt and Thomas defined DRY as knowledge — not code. AI generates duplicated knowledge at industrial scale. The discipline hasn't changed; the surface area has.
MCP servers promised to make your AI agent infinitely capable. They also made it slower, fatter, and easier to compromise. CLIs are quietly winning.
Every part of Claude Code maps to a part of a blender. Once you see it, you'll immediately know why your output came out wrong — and who to blame.
Building a Claude Code skill that encodes your own voice and engineering philosophy sounds like a neat AI trick. It turns out to be the most clarifying thing you can do with your expertise — because it forces you to name what you actually think.