“Research assistant” describes several products that work in very different ways. Some find and summarize web pages, some perform a series of searches, and others are general-purpose assistants with web access bolted on. Choosing one should start with the work we need to complete, not with a model leaderboard.

Start with the evidence

A citation-first answer is usually enough for a quick orientation. The goal is a concise map of the topic with links that make each important claim easy to inspect.

A decision memo or market analysis requires more. The tool needs to search more than once, compare sources, and identify where the evidence conflicts. A polished answer built on shallow sources can be worse than no answer because fluent prose makes weak evidence harder to notice.

For private documents, source boundaries matter more than the size of the public-web index. We need to understand what the product can access, whether existing permissions still apply, and how the provider uses the retrieved material.

Three types of research tool

Citation-first search

These products optimize for a fast answer with inline sources. They work well for learning the terminology, finding primary material, and getting oriented before deeper research begins.

Deep research workflows

Deep research products spend more time planning and running a series of searches. They can reduce the time required for a broad question, but the finished report still needs source-level review. More pages do not necessarily mean better evidence.

General assistants with web access

A general assistant is useful when research is one part of a larger task (for example, comparing public findings with internal documents and turning the result into a plan). However, browsing behavior and citation accuracy may be less consistent than they are in a product built specifically for search.

How we evaluate them

Use the same real question in each product, then compare five things:

  1. Source quality: Does the answer prefer primary and authoritative sources?
  2. Claim traceability: Can we connect each material claim to evidence that actually supports it?
  3. Coverage: Does the answer identify disagreement, uncertainty, and missing information?
  4. Control: Can we limit the search by domain, date, or source set?
  5. Workflow fit: Can we reuse the findings without spending time cleaning up the output?

The prose should not determine the winner. Open the citations. A product that writes an elegant summary of weak sources is a writing assistant, not a reliable research tool.

The default choice

Use citation-first search for frequent, time-boxed web research. Deep research makes sense when the breadth of the question justifies a longer wait and a more involved review. A general assistant is usually the better fit when private documents and the work that follows the research matter more than broad web coverage.

Regardless of the category, keep a human checkpoint before publishing the result or acting on a material claim. Research automation should shorten the path to evidence, not remove responsibility for evaluating it.