Field notes

How unattended agents fail in practice, with sources, and the check that catches each one.

Set a hard budget per scheduled agent run (Anthropic, OpenAI, OpenRouter)

Provider spend limits are monthly and account-wide. How to cap one scheduled agent run and one day, and pause the agent before the next run starts.

2026-10-05

Langfuse vs a watchdog: you don't need tracing to know your nightly agent is broken

Tracing tells you why a run went wrong. It cannot tell you the run never happened. Where Langfuse fits, where a watchdog fits, and how to run both.

2026-09-28

Our deploy job reported success for three days. The site stood still.

A brake counter, set -e and grep -c that exits 1 on zero matches killed our deploy three lines before it deployed. No error, no failed run, three days of a stale site, and one duration alert we ignored.

2026-09-16

n8n AI workflows: catch the failures your Error Workflow never sees

n8n's Error Workflow only fires when a node errors. Catch missed schedules, green runs that did nothing, and AI cost per run with two HTTP nodes.

2026-09-14

Cronitor for AI agents: what changes when the job is an LLM

A cron monitor tells you the job ran. An AI agent also needs a cost cap, loop detection and an outcome check. What to keep and what to add.

2026-09-07

Monitoring a headless claude -p job: hooks, exit codes, cost and alerts

A full tutorial for monitoring headless claude -p cron jobs with SessionStart, PostToolUse and Stop hooks: transcript cost, evidence checks and alerts.

2026-08-31

Prove what your AI agent did: an audit trail for unattended agents

A real, tamper-evident proof of one agent run: hashed record, public daily Merkle chain, Bitcoin anchor. How to verify it with a stdlib script, and what it cannot prove.

2026-08-26

My Claude Code cron ran up $1,800 in two nights. The watchdog that stops it at $2

Four documented runaway agent bills ($1,818, $6,000, $437, $150), why per-call metrics hide the loop, and per-run/per-day caps with rv.

2026-08-25

Claude Code Routine failed silently? How to know your scheduled agent actually ran

Green does not mean done. The 4 ways a scheduled Claude Code run fails with no error, and heartbeat plus evidence checks that catch each one.

2026-08-25

The dead man's switch for AI agents (and why a ping isn't enough)

A ping monitor knows a job ran. For an agent that is the least interesting fact. What a dead man's switch has to check when the job is an LLM.

2026-08-25

OpenClaw stuck in a polling loop: $150, a crash, and no alert. Detect it in 60 seconds

An OpenClaw agent called the same tool 1,535 times in two hours. Every call looked fine. How a retry-storm rule and a daily cap catch this before the bill.

2026-08-25

7 ways unattended AI agents fail silently, and one check for each

Seven ways a scheduled AI agent fails without an error, from a run that never starts to a run that does nothing, and the one detector that catches each.

2026-08-25

Coming up