Skip to slide
Chapter 15 · Deep Research Agents and How They Are Trained
119 / 142

CHAPTER 15 · Deep Research Agents and How They Are Trained · 9 / 10

Connecting to the bigger picture

Deep research agents are the multi-agent orchestration of Chapter 11 (MasterAgent, SubAgents, ReviewAgent), the memory mechanisms of Chapter 5, and the anti-hallucination contract of Chapters 10 and 13, all specialized for sourced output, plus a new training axis (SFT, RL/GRPO, CBR) that improves the model rather than the harness. The "harness matters more than the LLM" conclusion ties the whole guide together. The capstone borrows the review-pass idea as an optional fact-checking step.

← → arrow keys work too