Free playbooks in your inbox
Home / Series / AI Agent Operations Playbooks
Collection

AI Agent Operations Playbooks.

The two bills nobody budgets for: the security incident and the frontier model invoice. A threat model, a credential broker, a sandbox, plus a trace logger, an eval harness, and a router that proves a cheaper model keeps up on your own work.

2 titles · 10 free guides

Free guides
tutorial

Can DeepSeek Replace Claude on Your Code Reviews?

Replay 40 of your own code reviews against DeepSeek, score both models with the same pass test, and you get the sentence nobody in any thread can produce: it passed 22 where Claude passed 31, at 4.5% of the cost. Score Claude on the same traces or the number is worthless.

from: Stop Funding the Frontier

how-to

Route Changelogs to DeepSeek, Keep Claude on Code Review

Generate the routing table from your own scorecards and every rule carries the number that justifies it: the changelog moves at 0.6% of the cost with the pass rate held, and code review stays on Claude because no candidate cleared the bar. The rule that makes it safe is 5 words: no number, no route.

from: Stop Funding the Frontier

analysis

Claude Plans, DeepSeek Works: Is a Tiered Agent Actually Cheaper?

Split one agent so Claude plans and verifies while DeepSeek workers carry the volume, and you can price the bill per role instead of guessing which tier to tune. The saving is not the cheaper model, it is that Claude stopped reading things, which is also why the same split costs some teams 15x more.

from: Stop Funding the Frontier

reference

OpenRouter vs LiteLLM vs Ollama: Which One Should You Use?

All 3 answer how to reach a cheap model, not which one can do your job. Wire them behind one function, score the whole cheap tier on your own tasks by changing one string, and check the 3 things that disqualify a provider before price: context fit, data terms, and a free tier that stops dead at 50 requests a day.

from: Stop Funding the Frontier

analysis

How DeepSeek Can Fail 19 Times in 20 and Still Beat Claude on Cost

One line of arithmetic gives you the maximum failure rate a routed task can carry, and at a wide enough price gap that number is 95.5%. Narrow the gap and the same setup starts losing money at 3 misses in 10, which is why the model to reach for is the cheapest acceptable one, not the cheapest good one.

from: Stop Funding the Frontier

analysis 8 min read

The AI Agent Audit Log Lies. Here's How to Fix It

Your AI agent audit log says the human did it. Build distinct-identity attribution with an on-behalf-of trail so every agent action answers who really acted.

from: AI Agent Security: Lock Down Claude Code, MCP Servers & OpenClaw

tutorial 9 min read

Build a Credential Broker for Your AI Agent

Credential brokering stops your AI agent from holding a steal-able secret. Build a 30-line broker that mints scoped 5-minute tokens with a runnable test.

from: AI Agent Security: Lock Down Claude Code, MCP Servers & OpenClaw

reference 8 min read

MCP Server Security: Vet Any Server Before You Install

MCP server security for the people who install them: a four-gate vetting routine and the real package names, so a trojaned skill can't run as your agent.

from: AI Agent Security: Lock Down Claude Code, MCP Servers & OpenClaw

how-to 8 min read

How to Stop Prompt Injection in Your AI Agent

Prompt injection defense for AI agents: build a PreToolUse gate that blocks a live payload plus a memory-write guard that kills the slow poisoning attack.

from: AI Agent Security: Lock Down Claude Code, MCP Servers & OpenClaw

The series, in order

Read them in this order. Each one assumes the one before it.

01 Book
AI Agent Security: Lock Down Claude Code, MCP Servers & OpenClaw

AI Agent Security: Lock Down Claude Code, MCP Servers & OpenClaw

A threat model, scoped credentials, a sandbox, and a hardening runbook for every agent you ship

02 Book
Stop Funding the Frontier

Stop Funding the Frontier

A trace logger, an eval harness, and a router that proves a cheaper model keeps up on your own work

Want the buying order, not the shelf?

These titles sit on paths that tells you which three to buy first.