JJoeven

Curriculum

Agent Architectures

Simple agent loops from zero: six parts, ReAct, plan-and-execute, reflection, typed state, HITL, jobs, and when not to agent.

  1. 01

    Anatomy of an Agent

    Six named parts: assembler, model, parser, executor, memory, and stop. The loop is the product.

    20 min
  2. 02

    The Assembler Has a Budget

    The assembler decides what the model may see. Dumping the whole week plus 40 tool docs is how attention dies.

    19 min
  3. 03

    Parse, Then Fail Closed

    The model emits text. JSON plus a schema is a decision. Unknown names and broken JSON do not run.

    21 min
  4. 04

    Stop Conditions and Budgets

    Success, max steps, max dollars, max wall time, or handoff. Without stop, you bought a furnace.

    20 min
  5. 05

    ReAct: Thought, Action, Observation

    The classic loop: reason, call a tool, read the result, repeat. Implemented here with a fake model.

    22 min
  6. 06

    Observe Before You Finish

    A finish that ignores the last observation is a guess. The loop must read the world before it claims success.

    20 min
  7. 07

    Plan and Execute

    Write a short plan first, then run steps. Better for multi-hop work; worse when the world changes under you.

    21 min
  8. 08

    Plans Are Data, Not Poetry

    A plan is a list of typed steps with ids. Free-text paragraphs cannot be skipped, retried, or shown in a UI.

    19 min
  9. 09

    Reflection and Self-Critique

    A second pass that scores the draft against a checklist. It is not a second personality — it is a function.

    20 min
  10. 10

    Ground the Critic

    A critic that cannot see citations, tests, or policy will rubber-stamp vibes. Give it the same evidence the user will see.

    21 min
  11. 11

    Working Memory in the Loop

    Scratchpad, rolling summary, and retrieve-on-demand. The loop uses stores; it is not one vector soup.

    20 min
  12. 12

    Typed State Beats a Blob

    A dict with allowed keys is a contract. A giant string named state is how illegal tools sneak in.

    19 min
  13. 13

    State Machines for Agents

    Named states, allowed tools per state, and guards on transitions. Graphs you can draw beat loops you cannot.

    21 min
  14. 14

    Human in the Loop

    Pause before irreversible tools. The human sees a frozen payload. They do not become a free-text second model.

    22 min
  15. 15

    Freeze the Approval Payload

    Args at pause time are the contract. Resume must not pick up mutated dicts or extra keys from later thoughts.

    20 min
  16. 16

    Long-Running Agents

    A job with a store and a wakeup, not a request that holds a socket for an hour.

    21 min
  17. 17

    Checkpoints You Can Replay

    Save typed state after each step. Replay from a checkpoint instead of restarting the whole furnace.

    20 min
  18. 18

    Error Recovery

    Timeouts, retries with a cap, circuit breakers, and fail-closed. Recovery is policy, not vibes.

    21 min
  19. 19

    Computer Use in the Loop

    A screenshot grid is a last resort. Prefer an API. If you must click, bound the grid and never click pay.

    19 min
  20. 20

    Router, Specialist, Verifier

    Small agents with jobs beat one god-loop. Route first, specialize second, verify before the world changes.

    22 min
  21. 21

    Traces You Can Debug

    Every step logs assembler budget, raw model text, parse result, tool, observation, and stop reason.

    20 min
  22. 22

    Detect Tool Thrash

    The same tool with the same args three times is a bug. Stop, do not pay for a loop inside the loop.

    19 min
  23. 23

    Handoff Is a First-Class Stop

    Unknown, unsafe, or over-budget work goes to a human or another system with a packet, not a shrug.

    20 min
  24. 24

    When Not to Agent

    If a checklist, a form, or a search box will do, ship that. Agents are for branching work — and they are next to multiagent, not instead of a script.

    22 min
Start this track