Next Steps
Curriculum/128 lessons

The Next Steps — Agent Engineering in Depth

Lessons marked free are open to read in full. The rest unlock when you buy.

01 · Mechanics You Can Reason About

What is actually happening inside a request, at the level of detail that changes decisions: attention, position, tokenisation, sampling, reasoning budgets and caching.

1.1

What a forward pass actually does

Enough mechanism to stop guessing. Not a maths lesson — a working model of the thing you are delegating to.

6 min
free
1.2

Attention is the budget

The free course told you the smart zone exists. This is why it exists, and how to predict where yours is.

5 min
1.3

Position, recency and the middle

Where a fact sits in your context changes how strongly it lands. That is a lever, and most people never touch it.

5 min
1.4

Tokenisation, and why code is expensive

Your context budget is denominated in tokens, and code converts to tokens far worse than prose does.

5 min
1.5

Sampling, temperature and honest determinism

Where variance really comes from, which knobs help, and why reliability has to live outside the model.

5 min
1.6

Reasoning budgets: what effort actually buys

Extended thinking is real and measurable. It is also the most commonly misapplied setting available to you.

5 min
1.7

The prefix cache and the shape of your context

Providers cache the unchanged start of your context. Designing around that changes both your bill and your habits.

4 min
1.8

Tool calling under the hood

A tool call is generated text that your harness chose to execute. Everything odd about tool use follows from that.

5 min
1.9

Drift to the training distribution

Why agents write code that looks like a tutorial rather than like your repository, and what actually stops it.

6 min
1.10

What a better model actually changes

Which of your problems a model upgrade solves, which it makes worse, and how to decide whether to switch.

4 min

02 · Context Engineering

Treating the context window as something you design rather than something that fills: retrieval, session shapes, poisoning, memory systems, compaction and very large inputs.

2.1

The context budget as a design artefact

Decide what goes in the window before the session starts, the same way you would decide an interface before writing the class.

6 min
free
2.2

Retrieval that works

Why grep and the type system usually beat embeddings for code, and how to build a retrieval habit that scales.

5 min
2.3

Designing the opening message

The highest-leverage two hundred tokens you will write. A structure, and the failure each part prevents.

5 min
2.4

Five session shapes

Explore, decide, implement, verify, repair. Each needs different context, a different budget and a different ending.

5 min
2.5

Context poisoning

Some things in a window do not just waste space — they actively steer the agent wrong. Learn to spot them.

6 min
2.6

Memory systems, honestly

What persistent memory actually is, the two ways it fails, and how to get the benefit without the failure.

4 min
2.7

Compaction, deliberately

Automatic compaction is a lossy accident. The same operation performed on purpose is one of the better tools you have.

5 min
2.8

Very large inputs

A 40,000-line file, a 200 MB log, a database with 300 tables. Strategies that do not involve loading them.

6 min
2.9

Monorepos and multi-repo context

When the thing you are changing spans several packages or several repositories, context assembly becomes the hard part.

5 min
2.10

Instrumenting context

Three measurements that turn context management from a feel into a number, and what to do with each.

5 min

03 · Building Your Own Tooling

Stop working around your harness and start extending it: tools designed for agents, check commands as products, commands, skills, hooks, small MCP servers and scripted runs.

04 · Specification and Design

The discipline that decides whether unattended work is possible at all: interviewing yourself, acceptance criteria that bite, interface-first design, and specifying data, migrations and ambiguity.

05 · Orchestration and Multi-Agent Work

Running more than one agent without producing more than one codebase: subagents, parallel worktrees, orchestrator patterns, shared state, failure handling and knowing when to stop.

06 · Verification and Evaluation

The bottleneck, taken seriously: check ladders, property-based testing, a personal eval suite, reviewing at volume, detecting subtle breakage, and calibrating how much to trust.

07 · Legacy and Large Codebases

What changes when the code is old, large, undocumented or frightening: characterisation, seams, strangler patterns, undocumented behaviour and the politics of touching things nobody owns.

08 · Security, Safety and Supply Chain

The risks that come with delegating code and command execution: prompt injection, secrets, dependencies, permissions, data handling and reviewing security-sensitive diffs.

09 · Cost, Performance and Operations

Running this sustainably: where the money goes, latency and flow, model selection, budgets, failure modes in production and the operational habits that keep it working.

10 · Teams, Process and Craft

Making this work with other people: adoption, shared conventions, code review at volume, hiring and mentoring, and what happens to your own skill.

11 · Across the Stack

What actually differs by domain: frontend, data work, infrastructure, APIs, mobile and machine learning each break the general advice in their own specific way.

12 · Debugging and Incident Response

The skill agents help with least and cost most when done badly: hypothesis discipline, instrumentation, bisecting, production incidents, and the failures that only happen at scale.

13 · Architecture for Agent-Assisted Work

Design decisions reconsidered: what makes a codebase easy for agents, which classic trade-offs have shifted, and how to build systems that stay workable as generation gets cheaper.

14 · Research and Unfamiliar Territory

Using an agent to learn: evaluating libraries, reading unfamiliar code and papers, picking up a new language, and the specific ways confident wrongness bites hardest when you cannot check.

15 · Worked Projects

Six complete pieces of work, start to finish, with every decision and every mistake shown: a feature, a migration, an incident, a rescue, a greenfield service and a team rollout.

16 · Beyond Code

The rest of an engineer's job: writing, planning, estimating, technical decisions, runbooks, communicating with non-engineers, and knowing when not to use an agent at all.

Unlock all 128 lessons

One payment, lifetime access, every future revision included.

Buy for €149,00