Two-week cohortMon 28 September to Fri 9 October

The Agent-Era Engineer: From Prompting to Software Factories

Hand real work to an agent, and trust what comes back

Do you read the diffs your agent writes?

I did. Every one of them, for about a year.

I do not type code any more, and if you are reading this, neither do you. The agent writes it, runs the tests and opens the PR. For a while that felt like the whole shift. Then I noticed where my day had gone. I was reading. Four hundred lines here, nine hundred there, all afternoon, at a pace the agents had passed months earlier.

Two things that year taught me:

  1. Generation is solved. Not perfect, but solved in the way that matters: the code arrives faster than anyone can check it.
  2. So the bottleneck moved. It is no longer writing. It is knowing that what came back is right. Right now, that bottleneck is a person, and the person is you.

Once you see that, every other question about agentic coding gets easier. You stop asking how to get more out of the model. You start asking what has to be true before you trust its output without opening the diff.

Most engineers put their trust in one of two wrong places

The first is the agent’s word. It says “done, all tests pass,” and the PR merges. Sometimes the agent wrote the tests to pass. Sometimes they pass and the feature does not work in a browser, because nothing in the loop ever opened one.

The second is your own eyes. Nothing else in the pipeline has earned your trust, so you read everything. That works until the agents outrun you, and then it fails in the worst way: you skim, you approve, and you feel like you checked.

Both are the same mistake. The check lives in a person or in the author. It needs to live somewhere else.

First: decide what has to be true

Before any code exists, a change already has a definition of correct. It is in your head. The work is getting it out.

It starts at intake. Tasks pour in, and most of them can go straight to an agent. A person looks only at the ones that need a judgment call. That decision, made once and up front, is the first verifier you build. It continues through alignment, where an interview and multiple proposals surface what you wanted before an agent commits to a guess. It lands in the spec, where the question is not whether the spec is good but which verifier will say the work is done. Some tasks do not need a spec at all. There is a lesson on that too.

Second: build the check, and put it inside the loop

An instruction is hope. You wrote the rule in CLAUDE.md, the agent ignored it on Thursday, and you wrote it again in capitals. A verifier does not need to be remembered. It runs, it fails, the agent goes again. Prefer a verifier over an instruction, every time you can.

A real verifier touches reality: the deployed app, the real database shape, a browser that clicks through the flow, a Slack workspace that receives the message, a voice call that connects. A mock will pass anything. Building the environment where an agent can be checked against the real thing is most of the engineering in this course, and almost nobody teaches it.

Then the check goes inside the loop, not after it. A builder and a verifier, so the model grading the work is not the model that did it. Back pressure, so a failing check stops the agent and not you. Adversarial review from a different model with a reason to attack. The loop runs until it converges. You do not need the first draft to be right. You need a signal you trust to say when to stop.

Now: let it merge

Once the checks exist, the merge gate becomes a policy instead of a person. A copy fix with a passing run merges itself. A schema migration gets a verification environment, a test matrix and a hold for you. Verification scales with blast radius, and the skill is deciding that in advance so the agent knows too.

Your job changes shape. You stop reading diffs. You read outcomes, and you spend your attention on the changes that can actually hurt.

What I have built this way

I run 21 Dreams, a holding company that builds, scales and exits software products. AgentStack and Impello are two of the current ones, and they are the backbone of hundreds of companies. Every one of them is built and maintained the way this course teaches, and I have put 4,500+ hours into Claude Code and Codex doing it. Nobody at 21 Dreams types features by hand. Every change goes through agents, through the checks above, and through a merge gate that decides how much of it I need to see.

That is not a claim that it always goes smoothly. Verifiers go stale as a codebase moves under them. A reviewer model will agree with whatever it is shown unless you build it not to. Those failures are in the course, with lesson names like Verifiers Go Stale and Is Your Reviewer Any Good?, because you will meet them too.

Before and after

BeforeAfter
Reading every diffReading outcomes, and only the diffs that can hurt
“Done, all tests pass” as the signalA verifier that touched the real system as the signal
Rules in CLAUDE.md the agent forgetsChecks that fail the build and never forget
One reviewer model that agrees with the authorAdversarial review from a model that wants the bug
Every PR waits for youA merge gate that lets small changes through alone
Verifiers that quietly stopped meaning anythingScorers that grade the graders

If you already ship with an agent most weeks and want the reasoning, not a list of settings, this cohort is for you.

What you’ll learn

The lifecycle, in the order the work moves through it. Each stage hands to the next, and the last one hands back to the first.

The shift: generation is solved, so the bottleneck moved
  • Why the bottleneck keeps shifting, why verification is the whole game, and what convergence over perfection means for how you work.
Intake and alignment: what to hand over, and how far
  • The feed, triage, and the human gate. Then the interview, multiple proposals, and finding your unknowns before anyone writes a spec.
Specifications: should you spec at all?
  • What makes a good spec, the idea-to-spec skill, and choosing verifiers for a spec so it can be checked, not just read.
Context engineering
  • Anatomy of a node, signal to noise, progressive disclosure, the context layer. Why search isn't enough, and when to delete your README.
Loops and graphs
  • Primitives of a loop, where loops hide, the builder verifier pattern, worker loops, discovery loops and autoresearch.
  • The toolbox. /loop, /goal, the monitor tool, headless mode, routines and dynamic workflows.
Implementing, testing and verification
  • Back pressure, adversarial review, sycophantic attackers. TDD in the agent loop. Real verifiers touch reality, and where to set the bar.
Merging, maintenance and operating
  • The merge gate and where to auto-merge. Codebase entropy, migration seams, churn hotspots, dead code. Metrics, anomaly alerts, and the dogfooding agent.
Governance and the software factory
  • Identities, not keys. Scorers: agents that judge agents. Verifiers go stale. The supervisor loop and the self-improvement loop.

Learn directly from Ray

Ray Amjad

Taught 4,000+ Engineers. Ex-YC Technical Founder. Cambridge Physics.

I am a founder shipping real businesses with Claude Code & Codex. Since the tools launched I have put 4,500+ hours into them, building 21 Dreams, a holding company that builds, scales and exits software products, and the most in-depth Claude Code content on YouTube.

The workflows I teach come from production work, not a content calendar, which is why they tend to go mainstream months after I cover them.

This class is those 4,500+ hours concentrated: everything I know about shipping production-grade code with agents. My teaching style comes from my undergraduate Physics degree at Cambridge, so I teach fundamentals-first and pair every concept with the problem it solves, helping you build real intuition for a fast moving world.

Ray Amjad signatureRay Amjad signatureUniversity of CambridgeUniversity of CambridgeY Combinator

What’s included

4h
4 hours of live Q&A4 sessions a week: Tuesday and Thursday, each run twice for two time zones. Post your question in advance, upvote other people’s. Every one recorded.
25
The full class25 chapters and 236 pre-recorded lessons, about 22½ hours, from intake to the software factory. Each week’s chapters are assigned in order, and all of it is yours to keep after the cohort ends.
Δ
That quarter’s deltaWhat actually changed in the tools since the last cohort, and which of it is worth your attention.
Certificate of completionShare it with your employer or on LinkedIn.
30-day money-back guaranteeYour purchase is backed by the school’s 30-day money-back guarantee.

Cohort syllabus

236 pre-recorded lessons across 25 chapters • 4 live Q&A sessions a week • no live lectures

Before day 1 Prerequisites, recordedabout 6 h of recorded lessons
  • Welcome to the Class
  • The Ambiguity Line: Claude or Codex
  • Benefits of This Approach
  • Why Cloud Agents
  • Features of this Class
  • Creating Slack Workspace
  • Setting Up Bot Foundations
  • Connecting Claude Code
  • End to End Testing
  • Polishing the Bot
  • Connecting to GitHub
  • Dogfooding
  • MCP Servers
  • Playwright
  • Adding Database
  • CLAUDE md
  • Skills
  • Adding Memory
  • Adding Codex
  • Additional Tooling
  • Cron Jobs
  • This Chapter
  • On Call Agents
  • Task vs Ownership Delegation
  • Feedback Channels
  • Shallow vs Deep Modules
  • Leaky Abstractions
  • Blast Radius
  • The Map Is Not the Territory
  • Static Analysis
  • Long Context Failure
  • Getting Prompt Feedback
  • Build It Twice
  • Goal In, Strategy Out
  • Boxing the Agent In
  • Principles Over Rules
  • Dealing with Sycophancy
  • Coverage Through Stochastic Starting Points
  • Score Before You Spend
  • Point Fixes vs Architectural Fixes
  • Just Run It Again
  • Persona Vectors
  • Zero-Sum Attention
  • Instruction Following Limits
  • Build Small Merge Big
  • Infusing Lived Experience
  • The Pace-Layer Map
  • Bug Fixing Across Chats
  • Mixing Models & Modes
  • Using Reliable Packages
  • Agent Introspection
  • Intro to Context Engineering
  • The Skill
  • The One-Pattern Rule for Agents
  • Single Source of Truth
  • Grep Hygiene
  • One Feature Per Folder
  • Outside-Project Context
  • Architecture Decision Records
  • Knowledge Capture
  • Sentry's Quality Quarter
  • Gravitational Pull from Older Models
  • Clear Logging & Error Messages
  • Predictable Folder Names
  • Co-Locating Related Code
  • Fixing Leaky Abstractions
  • Clean Interfaces
Week 1 From intake to a running loop2 h live Q&A · about 7½ h of recorded lessons
  • The Big Ideas
  • Bottleneck Keeps Shifting
  • Generation is Solved
  • Introduction to Verification
  • Verification for Non-Dev Tasks
  • Convergence Over Perfection Thesis
  • The Feed
  • Triage
  • The Human Gate
  • The Interview
  • Using Fable for Interviews
  • Multiple Proposals
  • Multimodal Models for PRDs
  • Scoping APIs
  • Learn the Field First
  • Customized Terminology for Better Prompts
  • Build, Then Align
  • Unconstraining the Exploration Space
  • Find Your Unknowns
  • Glossaries
  • Stop Giving Claude Examples
  • Giving a Reference
  • Should You Spec?
  • What Makes a Good Spec
  • Idea to Spec Skill
  • Choosing Verifiers for Specs
  • Constraints Across Specs
Day2
Live Q&A: the shift, intake, alignment, specifications
Tue · live with Ray · 60 min · two time zones
  • The Prompt Isn't the Problem
  • Context Switching
  • Anatomy of a Node
  • Cognitive Inertia
  • Why Search Isn't Enough
  • Signal to Noise
  • Progressive Disclosure
  • The Context Layer
  • Example
  • Delete Your README.md
  • Different Orderings
  • Maintenance
  • Loops for Implementation
  • Test Time Compute
  • Intro
  • Primitives of a Loop
  • When a Loop is the Right Shape
  • Where Loops Hide
  • Your Role Over Time
  • Toolbox for Loops
  • /loop
  • Monitor Tool
  • Headless Mode & Background Workflows
  • Routines (aka Scheduled Tasks)
  • Memory for Routines (aka Scheduled Tasks)
  • Example: Databases
  • Dynamic Workflows
  • /goal
  • Writing Effective Goals
  • Codex Managing Codex
  • Builder Verifier Pattern
  • Example: Design Source of Truth
  • Designing a Task Lifecycle
  • Creating the Skill
  • Testing the Loop
  • My Daily Task Lifecycle
  • L4 Worker Loops
  • Don't Pre-Sequence the Backlog
  • L5 Discovery Loops
  • Autoresearch Overview
  • Autoresearch Technical Example
  • Ping Human
  • Autoresearch Non-Technical Example
Day4
Live Q&A: context engineering, loops and graphs
Thu · live with Ray · 60 min · two time zones
Week 2 From verification to the software factory2 h live Q&A · about 7½ h of recorded lessons
  • Implementation Notes & Decision Logs
  • Recursive Sub Planners
  • Back Pressure
  • Adversarial Review
  • Different Models for Reviews
  • Sycophantic Attackers
  • Customer Linters
  • Linters for Avoiding God Files
  • Quick Benchmarking
  • Understanding Agent Output
  • HTML Artifacts, Not Markdown
  • HTML Artefacts for Output
  • Interactive HTML Artifacts
  • Artifacts via Slack
  • Show Me Skill
  • Git Diffs & Mermaid Diagrams
  • Visualising Many Changes
  • TDD in the Agent Loop
  • Mutation Testing
  • Regression Testing
  • Backfill Testing
  • Why Verification is the Whole Game
  • Example: User Flows
  • Example: Measuring Completion
  • Real Verifiers Touch Reality
  • Validation is Adversarial
  • Prefer a Verifier Over an Instruction
  • Caveat to the Bitter Lesson
  • How to Verify Against the Plan
  • Setting Up Verification Environments
  • Test Matrices
  • Shims
  • Logging & Telemetry
  • Example: MCP Inspector Ladder
  • Runtime Context
  • Example: IP KVM & DevBoxes
  • Example: Namespace.so Verification
  • Example: Stress Testing Infra
  • Example: Slack Self-Verification
  • Example: E2E Voice Agents
  • Maintaining Verification Environments
  • Feature Maps As A Verification Source
  • Fuzz the Happy Path
  • Where to Set the Bar
  • Extreme Verification
  • Blast Radius Proportional Verification
  • The Merge Gate
  • Where to Auto Merge
  • Writing the PR
  • The Comprehension Gate
Day7
Live Q&A: implementing, testing, verification, merging
Tue · live with Ray · 60 min · two time zones

Lesson hours are an estimate. 120 of the 236 lessons are recorded, and the rest are sized at their average of about 6 minutes.

Student reviews

SimeonThe Course
Student

The Agentic Coding School doesn't just show you what buttons to press, it builds genuine understanding from the ground up. By the time you're two-thirds through, you're not just using the tool, you're thinking in it.

Art SmalleyThe Course
Student

This is not the typical get-rich-quick hustle program you see on social media. Ray works deep in the inner workings of the models and knows how to be productive with them. I am convinced he uses Claude Code as well as or better than most Anthropic engineers.

TylerThe Course
Student

I went from treating Claude Code like a chatbot to running it like a full development platform. The lessons on hooks, subagents, and CLAUDE.md files alone were worth it.

BernardThe Course
Student

I already had experience with Claude Code, and this still took me further. Seeing your everyday usage throughout the course is genuinely valuable, not just some vibe coding app. I learned a lot from the multi-agent plan execution and the spec-dev command.

DivyanshThe Course
Student

I am finding excellent value in the course. It serves as a vital reference library for me: whenever I hit a roadblock while pushing the tool's limits, I return to your material for practical tips. The biggest value is your experience as a builder.

DeividasThe Course
Student

This course is incredibly valuable for busy professionals who struggle to keep up with the fast-paced world of AI. He doesn't push unnecessary upsells but instead offers additional resources as genuine extras.

Frequently asked questions

Pre-recorded, so you can watch at whatever hour suits you. What is live is the Q&A: two sessions a week, each run twice so both time zones get one at a sane hour. Between the sessions there is a private Discord, so you are not holding a question until Tuesday: post it any day and I answer through my working hours.
No, but you do need to already be shipping with an agent. This starts from the assumption that you can drive the tool and are stuck on everything around it. If you are not there yet, start with the class instead. It builds the foundation this assumes you have.
The live hours are recorded, and the lessons are yours to work through whenever suits you. The one thing you cannot catch up on is asking your own question, which is why questions are submitted in advance rather than shouted into a call: you can post yours, and upvote someone else's, without being in the room.
Nothing. It is an email address. The cohort starts Mon 28 September, and the price and the enrolment window are announced to the waitlist before anywhere else, so the list is where you find out first, not where you get pushed.

For teams

Reimbursement
Get your company to pay

Everything your L&D needs: an email template, a receipt, and a certificate of completion.

Get reimbursed →
Private cohort
Run the two weeks for your org

The same lifecycle, on your codebase, on your schedule, with your team’s questions in the room.

Ask about company training →