← All cardsDOSSIER · Research assistant · SOLID · VERIFIED 2026-06-09
Elicit59SolidBenchmark pendingPaper discovery, evidence-table extraction, and systematic-review supportVerified 2026-06-09

Dossier · Research assistant

Elicit

Paper discovery, evidence-table extraction, and systematic-review support · last verified 2026-06-09

Solid
Research assistant
Elicit
59/100
ROLEPaper discovery, evidence-table extraction, and systematic-review support
Editorial fit
59
Source quality
52
Citation honesty
70
Privacy controls
46
Value for money
56
Speed
58
FREEBasic
$49/MOPro
$169/MOScale
SolidVerified 2026-06-09VP·METHOD

VERDICT HISTORY

  • Elicit is no longer just a paper-search assistant; it is positioning itself as an evidence-synthesis system for reports, systematic reviews, dynamic screening, extraction, and sentence-level citations over a very large scientific corpus. The richer data supports a higher-confidence card, but the recommendation still depends on benchmarked extraction accuracy against a hand-coded table.
  • Converted imported Notion research into a full flagship-ready dossier template with metrics, panes, pricing deck, scenarios, benchmark rows, and comparison slices.

EDITOR'S NOTE

Elicit should be judged by whether it helps a researcher build a better evidence table, not by whether it writes a convincing paragraph. Its danger is the same as its appeal: it makes a messy literature review look organized before the hard verification is finished.

AT A GLANCE

AI research assistant for searching, summarizing, chatting with papers, screening studies, extracting fields, generating research reports, and supporting systematic reviews across large scientific corpora and uploaded documents.

Role: Paper discovery, evidence-table extraction, and systematic-review supportCategory: Research assistantSynthesized

Elicit flagship-ready dossier: Paper discovery, evidence-table extraction, and systematic-review support.

PUBLIC FACTS · vendor & repo

List prices and pay-as-you-go entry points we can cite without running our own bench. Each tile links to a source when possible.

Research assistantSetVerdictPal card
synthesizedEvidenceQuality gate
6Sources checkedCard sources
2026-06-07Pricing checkedQuality gate

Limits & product surface

Non-price vendor claims — multipliers, caps, and API scope. Detailed matrices live in subscription and SDK sections below.

Primary surface
prosumer-saasCard identity
Modalities
text, dataCard identity
Workflow roles
[object Object], [object Object], [object Object]VerdictPal editorial
Alternatives tracked
Consensus, Scite, Semantic Scholar, Rayyan, Connected PapersVerdictPal comparison set

HOW IT WORKS · agent loop

The public positioning for this product — the loop we score against on VerdictPal.

Start with the wedge

Building structured evidence tables across many papers

Run the representative task

Extract PICO/design/outcome fields from 20 papers and compare against a hand-coded table.

Check the failure modes

Pricing jumps quickly from free to Pro/Scale

Compare before recommending

Compare against Consensus, Scite, Semantic Scholar before shipping advice.

METRIC LAB · 16 DIMENSIONS

Click a tile for the editorial note. Color follows score: coral, yellow, mint.

Shape

Avg 70 · 46–100

59Editorial fit

Weighted roll-up across 13 dimensions for Research assistant; pending desk verification if rescored from public facts.

By group

Fit80
  • Editorial fit59
  • Wedge task fit82
  • Feature depth100
Cost59
  • Free-tier utility67
  • Cost-to-value56
  • Opportunity cost55
Trust74
  • Source grounding70
  • Privacy posture46
  • Failure transparency100
  • Evidence strength52
  • Transparency100
Workflow67
  • Integration reach75
  • Setup friction58
  • Reliability70
  • Competitive position75
  • Data portability55

SEARCH MODES · 4 lenses

One input box, many retrieval postures. Filter by tier.

01Free

Workspace

Workspace path for Elicit — verify limits on the live product.

02Pro

Collaboration

Collaboration path for Elicit — verify limits on the live product.

03Pro

Automation

Automation path for Elicit — verify limits on the live product.

04Max

Admin & billing

Admin & billing path for Elicit — verify limits on the live product.

BENCHMARK LEDGER

Public rows are vendor or third-party claims we logged with a date. Desk rows are reserved for VerdictPal self-run results.

MeasureResultSource
Extract PICO/design/outcome fields from 20 papers and compare against a hand-coded tableexact-match accuracy, quote support, unsupported claims, missed caveats, hallucinated values, correction time.PlannedVerdictPal benchmark plan
Screen 200 records with known include/exclude labels and distractorssensitivity, specificity, false negatives, exclusion-reason quality, reviewer override friction.PlannedVerdictPal benchmark plan
Generate a research report from a controlled paper set; verify every claim against sentence-level citationscitation faithfulness, methods completeness, PRISMA trace, source coverage, overstated conclusion rate.PlannedVerdictPal benchmark plan
Model Basic, Pro, Scale, and Enterprise fit for a student thesis, policy brief, pharma SLR, and lab collaborationannual cost, report volume, table width, screening scale, collaboration needs, data terms.PlannedVerdictPal benchmark plan

PRICING DECK · checked 2026-06-09

Scale

$169

month

  • $2,028/yr annual billing
  • Verify live regional checkout before publishing procurement advice.

RAW · checked 2026-06-09: Basic free: 2 reports/mo, search/summaries/chat, 2 columns · Pro USD 49/user/mo billed annually: 5,000-paper systematic review workflow, 144 reports/reviews/yr, 20 columns, API · Scale USD 169/user/mo annual: collaboration, figures, 240 reports/reviews/yr, 30 columns · Enterprise custom.

RESEARCH LOG · desk notes

What we learned while building this dossier — not vendor copy.

Notion Card pipeline

Notion research pass

Elicit is no longer just a paper-search assistant; it is positioning itself as an evidence-synthesis system for reports, systematic reviews, dynamic screening, extraction, and sentence-level citations over a very large scientific corpus. The richer data supports a higher-confidence card, but the recommendation still depends on benchmarked extraction accuracy against a hand-coded table.

VerdictPal git

Flagship-ready structure

Converted imported research into metrics, panes, scenarios, benchmark rows, comparison notes, and pricing deck.

COMPETITIVE LENS

Where Elicit wins for cited research — and where a rival still belongs in the stack.

Elicit is stronger when building structured evidence tables across many papers; Consensus may still win for narrower fit, procurement, or specialist depth.

Elicit

  • Building structured evidence tables across many papers
  • Early-stage systematic review screening and extraction workflows

Consensus

  • You need a cheap casual search tool rather than a high-cost research workflow
  • Your review data, uploaded PDFs, or unpublished documents cannot leave local infrastructure without a DPA

UNDER THE HOOD

Vendors named on the product about page — useful for procurement and privacy reviews.

paper discovery
Elicit
claim extraction
Consensus
systematic review support
Scite

DEEP PANES · 6 LENSES

Editorial lenses only. Subscription and API pricing live in their own sections above.

TEST SCENARIOS · hands-on lab

How we exercised the product. Step through each run before you trust the scores.

Test run

Step 1 of 4

Start from the persona in best-for item 1.

BEST FOR

  • Building structured evidence tables across many papers
  • Early-stage systematic review screening and extraction workflows
  • Researchers who need cited summaries plus column-based extraction, not just a chatbot answer
  • Policy, pharma, medtech, and academic teams that value sentence-level citations, PRISMA-style method traces, and auditable extraction

AVOID IF

  • You need a cheap casual search tool rather than a high-cost research workflow
  • Your review data, uploaded PDFs, or unpublished documents cannot leave local infrastructure without a DPA
  • You expect PRISMA-grade results without human screening, extraction checks, and protocol review
  • Your main need is reference management or citation writing rather than evidence synthesis

STRENGTHS

  • Search across 138M+ papers
  • ClinicalTrials.gov source support
  • Paper chat with full-text access
  • Zotero import
  • Automated Reports
  • Systematic Review Workflow
  • Screening up to 5,000 papers on Pro and 40,000 on Enterprise
  • Extraction columns
  • Sentence-level citations
  • PRISMA/methods/search-strategy report outputs
  • API access

WEAKNESSES

  • Pricing jumps quickly from free to Pro/Scale
  • Structured tables and reports still require human validation
  • Uploaded papers/Zotero imports create cloud privacy boundary
  • Corpus and full-text coverage vary by discipline
  • Vendor accuracy claims need independent reproduction
  • Basic plan limits can distort real review workflows

HOW IT COULD IMPROVE

  • Institutional policy friction could be reduced with a local/on-prem deployment option, or at minimum with clear data-handling documentation that compliance teams can review without an NDA.
  • Coverage gaps could be narrowed by expanding the corpus beyond the current strong disciplines, or by surfacing coverage limitations explicitly so the reader knows where not to trust the tool.
  • Privacy posture scores 46/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Evidence strength scores 52/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.

EDITORIAL EVIDENCE · 4 ENTRIES

TASK

Extract PICO/design/outcome fields from 20 papers and compare against a hand-coded table.

Pending.

exact-match accuracy, quote support, unsupported claims, missed caveats, hallucinated values, correction time.

TASK

Screen 200 records with known include/exclude labels and distractors.

Pending.

sensitivity, specificity, false negatives, exclusion-reason quality, reviewer override friction.

TASK

Generate a research report from a controlled paper set; verify every claim against sentence-level citations.

Pending.

citation faithfulness, methods completeness, PRISMA trace, source coverage, overstated conclusion rate.

TASK

Model Basic, Pro, Scale, and Enterprise fit for a student thesis, policy brief, pharma SLR, and lab collaboration.

Pending.

annual cost, report volume, table width, screening scale, collaboration needs, data terms.

This tool is in the VerdictPal Citation Fidelity protocol — frozen question bank, dual-reviewer grading, results still pending. Citation Fidelity v0.1 →

PRIVACY DEEP-DIVE · checked 2026-06-09

Elicit collects account, device/usage, and service-created research data. Uploaded papers, Zotero imports, BYO documents, screening criteria, extraction tables, and report prompts are cloud-processed. Enterprise pricing states no training on customer data by default; do not assume that boundary for all lower tiers without terms review.

Training on your dataOff by default
EU data residencyEnterprise tier only
SOC 2 attestationNot documented
Local-first by defaultCloud only

WORKFLOW ROLES

How this tool fits into a composed research stack:

paper discoveryclaim extractionsystematic review support

QUALITY GATE · SOLID

Evidence ready
Benchmark pending
  • score-revision 2026-06-09: differentiated from public facts; pending desk verification
  • Flagship-quality promotion 2026-06-04: source-backed dossier promoted; benchmark and public recommendation remain locked until evidence packet is complete.
  • Next action: Run 20-paper extraction benchmark and 200-record screening recall test; verify every report claim against sentence-level citations.
  • Has API: true
  • Public recommendation: false
  • evidence-override: Notion source-of-truth sync; evidence remains source-backed until desk benchmark or hands-on pass.
  • Flagship-ready structure generated from Notion deep research and imported git fields on 2026-06-09.

Related dossiers

Browse the atlas

The Pack · Editorial newsletter

New cards in your inbox. Free.

One short email when a card ships or changes status. No tracking, no third-party analytics. Unsubscribe in one click.