Skip to slide
Chapter 15 · Deep Research Agents and How They Are Trained
116 / 142

CHAPTER 15 · Deep Research Agents and How They Are Trained · 6 / 10

Why it matters

DR agents are where the harness patterns meet the training methods, and they preview where coding agents are heading: planned, multi-agent, self-verifying, and increasingly trained rather than only prompted. Even if you never train a model, knowing the vocabulary (SFT, RL, GRPO, CBR) lets you read the field and understand why an agent behaves as it does. And the architecture (MasterAgent, SubAgents, ReviewAgent) plus the fact-checking loop are directly reusable in any agent that must produce trustworthy, sourced output.

← → arrow keys work too