Build with Claude
A course-style series on running Claude Code and agentic pipelines in production, built entirely from incidents on this site's own infrastructure — what broke, why it was hard to see, and the actual fix.
Most Claude Code content is a walkthrough of the happy path. This is the other kind: specific things that broke while running Claude Code and agentic pipelines on this site's own infrastructure, why the failure was hard to see before it was fixed, and what actually changed. Each lesson names the mechanism, not just the moral.
Cost & Infrastructure Economics
What it actually costs to run Claude Code day to day and on a schedule — and the specific changes that cut real spend without cutting quality.
- 1 The Cheapest Way to Run Scheduled Claude Jobs Isn't the API →
- 2 The Env Var That Secretly Tripled Our AI Coding Bill →
- 3 When Prompt Caching Costs You More Than It Saves →
- 4 Why Your Spend Limit Doesn't Survive a Fresh CI Checkout →
- 5 How Five Free Workflows Added Up to a Real Bill Coming soon
- 6 Stop Letting Your AI Agent Reroll From Scratch After Every Rejection Coming soon
Silent Failures
The absence of an error is not the same claim as "it worked." Seven incidents where a pipeline looked healthy while doing nothing, or the wrong thing.
- 7 The Page That Loaded Fine and Showed Nothing Coming soon
- 8 The Eighteen Tests That Never Actually Ran Coming soon
- 9 We Built a Drift Detector That Could Never Detect Drift Coming soon
- 10 Your Alerts Need Their Own Alarm Coming soon
- 11 The Outage Where Every Health Check Said Everything Was Fine Coming soon
- 12 A Feature Being Off Looks Exactly Like a Feature Being Broken Coming soon
- 13 One in Nine of Our Articles Was Cut Off Mid-Sentence Coming soon
Credentials & Secrets
Agentic workflows change how credentials get created, checked, and leaked. Four incidents that weren't about a stolen key — they were about a check that lied.
- 14 A Redaction Script Hid Every Secret Except One Coming soon
- 15 Why Our Health Check Said a Working Key Was Dead Coming soon
- 16 The Error Message That Sent Us In the Wrong Direction Coming soon
- 17 The Day Our Three Secret Stores Disagreed Coming soon
Running Claude Unattended
Scheduling agentic work and coordinating more than one agent surfaces failure modes interactive sessions never hit.
- 18 Our Local Cron Jobs Kept Missing Their Runs Coming soon
- 19 The Scheduled Jobs Nobody Told the Scheduler About Coming soon
- 20 A Green Checkmark Doesn't Mean the Work Got Done Coming soon
- 21 What Happens When Two AI Agents Race to Log the Same Event Coming soon
- 22 The Auto-Merge Setting That Doesn't Actually Protect Anything Coming soon
Deploy, CI & Git
Shipping what an agent builds hits its own class of problems — most of them about ordering, not code.
- 23 The Deploy Command That Broke Three Different Ways in One Month Coming soon
- 24 A Squash Merge That Changed Nothing and Broke Everything Coming soon
- 25 Our Security Scanner Got Cancelled by the Step Before It Coming soon
- 26 The Webhook That Was Faster Than Its Own Database Coming soon
- 27 One Leftover File Broke Every CI Run for a Week Coming soon
Running Claude Code in production is a program, not a prompt
If your team is scaling from individual Claude Code use to agentic pipelines running unattended, most of the failure modes above show up eventually. Happy to talk through what we've built.
Work with me →