Forward-deployed / Learning zone
AI agentsa Generative AI module
A standalone module

AI agents for the product leader

The loop behind every agent, how much autonomy a task actually needs, what keeps it reliable across many steps, and the economics that decide whether it's worth building at all.

3 lessons+ recapknowledge graphdiagrams included

An agent is what you get when you give a model a goal, a set of tools, and permission to loop: gather context, decide an action, take it, look at what happened, and repeat until the goal is met or the budget runs out. That loop is the single idea behind every "AI that does things" pitch, and it is also where an AI feature's economics and its failure modes both actually live — not in the model's raw intelligence, but in how the loop is scoped, watched, and stopped.

A note on scope. This topic has the deepest existing coverage of anything in this family. Agentic AI for the AI PM already develops the loop, the autonomy spectrum, planning and reasoning patterns, reliability and the compounding-error math, and the economics of when an agent actually pays for itself, in full depth and in the same product-leader voice this family uses. Two of the planned lessons for this module — tools, and memory across a run — are also already the dedicated subject of two sibling modules in this family, Tool calling and Memory & context. Re-deriving any of that here would only restate it three times over. This module exists to be the compact front door under the name people actually search for — AI agents — and to send every reader past it into the depth that already exists. That is why it is three lessons, not seven: the honest amount of genuinely new ground, once the sibling modules and the deeper agentic-ai track are accounted for, is three lessons' worth.

The knowledge graph

AI agents

A decision funnel · not a feature list

Three questions to ask before writing a single line of agent code.

Lesson 1 · What it is
The loop & the autonomy dial

Every agent is one loop plus a setting on one dial — how much your code decides, versus the model.

The loop

Gather · decide · act · observe

Autonomy dial

How much the model decides for itself

Lesson 2 · How it runs
Planning · reliability · the two things that decide it works

Not how smart the model is — how the loop compounds when it's chained.

Planning & reasoning

The patterns · decompose · reflect · retry

Reliability across many steps

Where 95% × 20 = 36% shows up

Lesson 3 · Whether to build it
Stakes · verifiability · volume

Three questions in sequence that catch most of the bad agent bets before they ship.

Cheap to check?

If not — assist, don't automate

Mistake reversible?

If not — human approval gate

Volume enough?

If not — one-off is cheaper

All three · yes

The sweet spot for autonomy

Often the answer
A fixed workflow

Cheaper, more predictable, easier to debug — pick this whenever the autonomy isn't earning its cost.

Read it as a decision funnel, not a feature list. What it is: the loop is simple, and the first real design choice is how much autonomy the task actually needs. How it runs: once it's looping, the two things that determine whether it works are how it reasons and how well it survives many steps in a row without its errors compounding. Whether to build it: the loop above is often the wrong answer, and knowing that before you build it is the highest-leverage call in this module.

The lessons

Each lesson pairs the product framing with a 🎯 For the product leader briefing — why it matters, the decision it changes, the question to ask your team, and the risk if ignored — plus a diagram. For the engineering depth behind every mechanic mentioned here, follow the spokes into Agentic AI for the AI PM, and for tools and memory specifically, into Tool calling and Memory & context.

📌 Close out the module: Recap & real-world examples.

The lessons

01

What an agent is, and how much autonomy it needs

Read lesson →
02

Planning, reasoning & reliability across a run

Read lesson →
03

When not to build an agent

Read lesson →
📌

Recap & real-world examples

Read recap →