Skip to content

Glossary

Deep research

Deep research is a mode that spends minutes rather than seconds, reading across many sources and returning a written brief with references, so the thing being bought is breadth of reading rather than a cleverer answer.

In plain terms

Asking a question and getting a short report back instead of a paragraph. It goes away for a few minutes, reads a lot of pages, and returns something written up with links to where each part came from. The extra time is spent on reading more widely, not on thinking harder, which is a smaller difference than it sounds and a useful one.

01

Why it matters

Because it changes what the output is for. A quick answer is something you read and act on, while a brief with references is something you check, and those are different working relationships with the same tool. Treating the second like the first is how a plausible summary of thin sources ends up inside a decision nobody revisited.

02

How it works

The defining change is spend rather than capability. The system issues many searches, opens many pages and works through them in sequence instead of answering from what it already holds, which costs minutes and real money per run. Nothing about the underlying reasoning improves; there is simply far more material in front of it when it writes.

The corpus it can reach decides what it can possibly find, and the corpus differs sharply between products. Some read the open web, some are restricted to academic literature, and some answer only from documents you supplied and will not go looking at all. A question about market pricing put to an academic tool returns a careful, well-referenced answer about the wrong thing.

References are the deliverable, which is the part most readers under-use. Each claim traces to somewhere, so the brief is checkable in a way an ordinary answer is not, and the check takes a click rather than a re-run. Skipping it converts the whole advantage back into a confident paragraph you have no way to test.

The characteristic failure is a good brief built on thin sources. Length and structure make the output feel authoritative, and neither is evidence about what it read, so a topic with little serious material published on it produces something that looks exactly like a topic with a great deal. The tell is in the references rather than in the prose, and it is visible in seconds if anybody looks.

The written brief has started arriving as other shapes. Several products now return slides, spreadsheets or a page rather than prose, which is convenient and quietly raises the stakes, because a finished-looking artefact invites less checking than a wall of text does. The more presentable the output, the more deliberate the verification step needs to be.

What the extra minutes change

What the extra minutes changeThe reason this mode gets over-trusted is that every visible signal moves in the same direction while only half the underlying quality does. The answer arrives longer, structured under headings, carrying references, after a wait that itself implies effort. All four of those are real, and none of them is evidence that the material behind it was worth reading. A question with a thin published record produces the same artefact as a question with a rich one, because the shape comes from the system's writing habits rather than from what it found. The habit that fixes this costs almost nothing: open two or three references before reading the summary, and specifically the ones under whichever claim the decision will rest on. If those are solid the rest usually is, and if they are a vendor's own marketing page dressed as a finding, that is worth knowing before the brief is quoted in a meeting.Gets betterHow much was read.How many sources are compared.Whether an obscure itemsurfaces.How checkable the result is.Stays the sameHow well it reasons.Whether a source was any good.Whether a claim was misread.How confident the prose sounds.The right-hand column isunchanged and the output looksfar more authoritative anyway,which is the whole of the risk.Length reads as rigour, and itis a property of the writingrather than of the reading.
The reason this mode gets over-trusted is that every visible signal moves in the same direction while only half the underlying quality does. The answer arrives longer, structured under headings, carrying references, after a wait that itself implies effort. All four of those are real, and none of them is evidence that the material behind it was worth reading. A question with a thin published record produces the same artefact as a question with a rich one, because the shape comes from the system's writing habits rather than from what it found. The habit that fixes this costs almost nothing: open two or three references before reading the summary, and specifically the ones under whichever claim the decision will rest on. If those are solid the rest usually is, and if they are a vendor's own marketing page dressed as a finding, that is worth knowing before the brief is quoted in a meeting.
03

Seen in the wild

  • An answer product whose research mode reads autonomously across many sources and returns a structured report rather than a paragraph.

    Perplexity
  • A general assistant offering a documented-report mode alongside its ordinary chat, for work that needs sources attached.

    ChatGPT
  • A tool restricted to material you upload, which answers only from that corpus and will not wander to the open web unless asked.

    NotebookLM
04

Common misconceptions

People assume

It is a cleverer version of the same thing.

In fact

It is a longer version, and that distinction decides what to expect. The same reasoning is applied to much more material, so it finds things a quick answer would have missed and reasons about them no better. Problems that were hard because they needed judgement stay hard; problems that were hard because they needed reading get easier.

People assume

The references mean the brief is verified.

In fact

They mean it is checkable, which is a different and more useful thing. A source can be misread, thin, outdated or simply not say what the sentence citing it claims, and none of that is visible from the prose. The value arrives when somebody clicks, and a brief nobody opens the references on is just a longer answer.

People assume

A longer run gives a better answer.

In fact

It gives a wider one, and past a point that is not the same. Once the genuinely relevant material has been read, more time buys weaker sources and more restating, so the useful length tracks how much has actually been published on the question. Thin topics do not get better answers from longer runs.

05

Telling them apart

Deep research vs Answer engine

Deep research

Minutes, many sources, a written brief you check.

Answer engine

Seconds, a direct answer with its sources attached.

Same instinct at two speeds, and the deliverable differs: one produces something to read now, the other something to review before it is used.

06

Questions

When is it worth the wait?
When the question needs breadth rather than judgement, which is a narrower set than it first appears. Comparing what several sources say, surveying a landscape or assembling background all reward the extra reading. A question whose difficulty is deciding between two defensible positions gets a longer answer and no more help than a quick one would have given.
How do I tell a good brief from a confident one?
Open the references, and do it before reading the prose rather than after. Thin topics produce briefs that read exactly like well-supported ones, because length and structure come from the writing rather than from the material. Two minutes spent on where the claims came from tells you more than a careful read of the summary.
Does the tool matter, or is it the same feature everywhere?
The corpus differs and that decides what any of them can find. Open-web tools, academic tools and tools restricted to your own uploaded material will answer the same question differently and correctly within their own scope. Matching the corpus to the question matters more than which product has the better research mode.
Can I trust it for something load-bearing?
Treat it as a well-organised first pass by somebody who read quickly, which is genuinely valuable and is not the same as verified. For anything a decision rests on, the references are the work: check the ones the argument depends on rather than all of them, since load-bearing claims are usually a small share of the brief.
07

Key takeaways

  • The mode buys reading time, not better reasoning.
  • The corpus it can reach decides what it can possibly find.
  • References make it checkable; nobody checking makes it a longer answer.
  • A thin topic produces a brief that reads exactly like a well-supported one.
  • The more finished the output looks, the more deliberate the check should be.
09

Tools that use this

  • Perplexity

    A research mode reading across many sources into a report.

  • ChatGPT

    A documented-report mode alongside ordinary chat.

  • NotebookLM

    Restricted to your uploads, answering only from that corpus.

Last checked July 2026

All glossary terms