Skip to main content
Back to Course 102
4/13
Unit 4 of 13
Unit 03 — Phase 01

Research Like a Team of Ten

Configure the Machine

Listen to this unit

Read-aloud narrates the unit in a natural voice, paragraph by paragraph. It needs an account.

In short

Delegate the reading, never the believing. Research agents have superhuman stamina and mediocre judgment; belief stays with the person who acts. Commission with a brief, not a question: scope, decision, prefer, avoid, shape, flags. An autonomous agent multiplies whatever steering you give it, including none.

Socratic Mode

The Hook

There's a kind of AI product that doesn't answer your question in three seconds. You ask, and it disappears for ten minutes. While it's gone, it runs dozens of searches, opens hundreds of pages, follows leads, discards dead ends, and comes back with a multi-page report, structured, sourced, and footnoted.

They're called research agents, and every major AI provider now ships one. Used well, they compress a day of reading into a coffee break. Used badly, they produce something more dangerous than ignorance: a beautifully formatted report you didn't check, wearing the costume of diligence.

The difference between those two outcomes isn't the agent. It's whether the person who commissioned it understands the one rule of delegation this unit is built on.

Video — The HandoffWatch on YouTube

The Core Concept

The mental model: delegate the reading, never the believing.

A research agent is a reading machine of superhuman stamina and mediocre judgment. It will genuinely cover more ground in ten minutes than you could in a day. What it won't reliably do is weigh source quality the way you would, notice when two sources are copying the same original error, or care whether a claim is load-bearing for your decision. Those are acts of believing, and believing is the part that stays human, because you're the one who acts on the result.

The landscape as of August 2026: deep research modes in ChatGPT, Gemini, and Claude (each provider names it slightly differently) send an off to investigate for minutes at a time and return long, cited reports. Perplexity (perplexity.ai) is a search-first assistant that answers with citations by default and has a deeper research mode of its own. And NotebookLM (notebooklm.google.com), your grounded workspace from Unit 01, is where research goes after it's gathered. Free paths exist across all of these; deep research modes on free tiers are usually capped at a few runs per month, which is enough for this course.

Notice what just happened in that last sentence: the tools formed a chain. That chain is this course's central pattern, so it gets a name. The Handoff is the moment one tool's output becomes the next tool's input, deliberately, with you deciding what crosses. From this unit forward, every Live Demo ends with one. Stringing handoffs together is how single tools become systems, and it's the road that ends, in Unit 08, with systems that run on their own.

The research version of the chain is the verified research relay, and it's this unit's artifact:

  1. Commission a research agent with a brief, not a question (breadth).
  2. Hand off the report plus its best sources into a grounded notebook (depth).
  3. Interrogate the notebook, where every answer cites your actual sources.
  4. Verify the load-bearing claims yourself, at the original source (belief).

It's how a good newsroom works. Stringers and wire services gather everything (breadth). An editor pulls the credible material into one working file (the handoff). The journalist works that file hard (depth). And before publication, the load-bearing facts get checked against the primary source, by a human whose name goes on the story (belief). Nobody prints the wire feed raw. The relay exists because gathering and believing are different jobs.

Explained your way: the AI rewrites this idea around something you already know.

Why does the commissioning step matter so much? Because a research agent inherits every ambiguity in your request and multiplies it across a hundred pages. "Tell me about creatine" produces a generic tour. A brief produces an investigation: scope, the decision this feeds, sources to prefer, sources to avoid, output shape, and what to flag as uncertain. You already know this pattern; it's the instruction block from Unit 01, applied to a single mission instead of a standing workspace. Draft one properly before you spend a research run on it:

Context Engineering Lab
Unit 03 · Research Like a Team of Ten

Edit the prompt on the left. Click "Run" to see how the AI output changes. Try adding context — who you are, what you need, what good looks like.

AI output
Dear Hiring Manager, I am writing to express my interest in the position at your company. I am a motivated individual with strong skills and experience that make me a great fit for this role. I look forward to hearing from you. Sincerely, [Your Name]
Running a prompt against a live model needs an account.

The legal profession ran this unit's experiment at scale, involuntarily. Since 2023, courts around the world have sanctioned lawyers for filing documents containing AI-fabricated citations, and public trackers of such incidents passed several hundred cases in US courts alone by 2025, with new ones still landing in 2026. The pattern is always the same: fluent, formatted, confident research output, submitted by a professional who delegated the believing.

Modern research agents make this failure quieter, not rarer. Their citations are usually real now. But real citations can still point to weak sources, misread the source they point to, or lean on one blog post that three other cited pages were themselves quoting. The costume of diligence got better. The rule didn't change: the human who acts on the report checks the claims the action rests on.

A forty-source report feels rigorous. But agents count sources; they don't weigh them. Five of those forty may be reprints of the same press release. The decisive claim may rest on the weakest link in the bibliography. And a claim repeated across many pages of the internet gets found more often, which means an agent can mistake popularity for truth exactly the way a search engine does.

So don't audit the whole report; that's delegating the reading back to yourself. Find the two or three claims your decision actually rests on, and chase those to their original sources. Depth on what matters beats breadth of suspicion.

Knowledge check
Two students research the same purchase decision. Amal types "best budget laptops" into a research agent. Zain gives it: his budget, his three uses, a request to prefer manufacturer specs and independent reviews over affiliate listicles, a comparison-table output shape, and an instruction to flag anything uncertain. Why will Zain's report be dramatically more useful?
Knowledge checks save to your account.

Live Demo

Free path: Gemini's and ChatGPT's deep research have limited free runs, Perplexity's core search is free, and NotebookLM is free. One free deep-research run is all this demo needs; Perplexity substitutes if you've used yours.

Step 1, commission the Studio's brief. The Studio needs a weekly trend report to draft posts from. Here's its commissioning brief; read it as the worked example of the shape:

Prompt
Research current trends in neighborhood coffee shop marketing for the coming month. Scope: small independent shops, not chains. Prefer: industry publications, platform-published creator guidance, and named case studies. Avoid: listicles with affiliate links and content older than 12 months unless it's foundational. Output: five trends, each with what it is, one real example, and one way a small shop could act on it this week. Flag anything you're uncertain about instead of smoothing it over.

Step 2, commission your own. Pick a real question from one of your Unit 01 domains, something you'd genuinely spend an afternoon researching. Write your brief with the same anatomy (scope, decision, prefer, avoid, shape, flags) and launch a deep research run.

Step 3, The Handoff. When the report returns, don't stop at reading it. Export or copy it into a new NotebookLM notebook, and add two or three of the report's most important original sources alongside it. The agent's output just became your workspace's input. That move is the Handoff, and you'll make it in every unit from now on.

Step 4, interrogate. Ask the notebook the questions that matter for your decision. Every answer now cites your actual sources, and clicking a citation takes you to the passage. Ask at least one question designed to stress the report: "What do these sources disagree about?"

Step 5, verify the load-bearing three. Name the three claims your decision rests on. Chase each to its primary source, outside the notebook if needed. Grade each: holds, holds with caveats, or doesn't hold. If any claim fails, note what the agent did wrong: bad source, misread source, or overconfident synthesis.

The report reader
Asks a bare question, receives forty formatted pages, reads the summary, acts on it. Feels thorough because the report looks thorough. Has outsourced breadth, depth, and belief in one click, and will discover the weak claim at the worst possible time.
The relay operator
Commissions with a brief, hands the report into a grounded notebook with its key sources, interrogates it against the real decision, and personally verifies the three claims that carry weight. Ten extra minutes. The belief stays home.

Operator Moves

Commission with a brief, never a question. Six lines: scope, the decision this feeds, prefer, avoid, output shape, flag uncertainty. Save the skeleton in your Unit 01 workspace and reuse it; a research run is too expensive, in time and in free-tier quota, to spend on a guess.

The three-claim spot check. After any research output, name the claims the decision rests on (there are almost never more than three) and chase them to primary sources. This is the whole verification budget, spent where it pays.

Never cite what you haven't opened. If a claim from AI research is going into your essay, your presentation, or your decision, you open the original source first. No exceptions. This single rule is what separates you from every cautionary tale in the case study above.

Why This Matters

This unit is where you stop being one person. A commissioned agent plus a grounded notebook plus a verification pass genuinely covers what a small research team covers, and it does it on free tiers, on a phone. The students and professionals who learn to commission and verify get compounding returns on every question they'll ever ask.

It's also where the course's central pattern locked in. The Handoff you made in Step 3 is the same move you'll make when a terminal agent's work flows into your files (Unit 04), when a skill packages a workflow for reuse (Unit 05), and when a graph strings the whole chain together to run overnight (Unit 08). Systems are just handoffs made permanent. You now know what crosses each gap, because you carried it across by hand.

And the rule scales with the power. The more capable agents get, the more tempting it becomes to delegate the believing, and the more expensive that delegation gets. The operators who thrive are the ones whose verification discipline grows in proportion to their delegation.

Knowledge check
A research agent's report says a supplement improves sleep "according to multiple studies," citing four sources. Before acting on it, the highest-value check is:
Knowledge checks save to your account.

The Challenge

The Verified Research Relay

45 minutesHands-on

Run the full relay on a question from your own life, and keep the artifact.

  1. Pick a real question from one of your Unit 01 domains, one where being wrong would actually cost you something (a purchase, a study choice, a health or training decision, a project direction).
  2. Write the commissioning brief with all six parts: scope, the decision it feeds, prefer, avoid, output shape, flag uncertainty. Save it to the relevant workspace.
  3. Run the research with any deep research tool (free path above) and complete The Handoff: report plus two or three key original sources into a grounded notebook.
  4. Interrogate the notebook with at least three decision-relevant questions, including "what do these sources disagree about?"
  5. Verify the load-bearing claims: name up to three, chase each to its primary source, and grade each: holds, holds with caveats, doesn't hold.
  6. Write the verdict: five sentences. What will you actually do, which verified claims support it, and what did the agent get wrong or overstate?
Success criteria: a saved brief, a grounded notebook containing the relay, three graded claims with primary sources named, and a decision you can defend claim by claim without opening the AI's report again.
Submitting your work needs an account.

Key Takeaways

  1. 1Delegate the reading, never the believing. Research agents have superhuman stamina and mediocre judgment; belief stays with the person who acts.
  2. 2Commission with a brief, not a question: scope, decision, prefer, avoid, shape, flags. An autonomous agent multiplies whatever steering you give it, including none.
  3. 3The Handoff is the course's central pattern: one tool's output becomes the next tool's input, with you deciding what crosses. Breadth agent, grounded depth, human belief.
  4. 4Verify by triage: the two or three load-bearing claims, chased to primary sources. Never cite what you haven't opened.

The Rabbit Hole

Type: Article Title: Introducing deep research, OpenAI URL: https://openai.com/index/introducing-deep-research/ Description: The launch essay for the product category this unit tames, including the makers' own framing of what it's for and where it fails. Read it as an operator: notice which claims are about gathering, and which quietly reach toward believing.

Explore Further

TypeTitleURLDescription
ArticleOpenAI, "Introducing deep research"openai.com/index/introducing-d…The defining product launch for autonomous research agents
ArticleGoogle, Gemini Deep Research overviewgemini.google/overview/deep-resea…Google's version, with free-tier access notes
DocsAnthropic, Claude Research helpsupport.claude.com/en/articles/1108886…How Claude's multi-step research feature works and which plans include it
ToolPerplexityperplexity.aiSearch-first assistant with citations by default; free tier
ToolNotebookLM (Gemini Notebook)notebooklm.google.comThe grounded destination for every research handoff
Case trackerAI Hallucination Cases databasedamiencharlotin.com/hallucinationsThe running public tally of fabricated citations in court filings
ArticleCalMatters, AI-fabricated citations in US courts (2025)calmatters.org/economy/technology/…How delegated believing plays out at professional stakes

Last updated: August 9, 2026. Product names in this unit shift; the relay and the verification discipline don't.

Track your progress

Marking a unit complete, the Prove It check and your place in the course all need an account. The reading stays free.