← All toolsDOSSIER · General AI assistant · SOLID · VERIFIED 2026-06-09
ChatGPT68SolidBenchmark pendingGeneral-purpose assistant baseline and multimodal draft criticVerified 2026-06-09

Dossier · General AI assistant

ChatGPT

General-purpose assistant baseline and multimodal draft critic · last verified 2026-06-09

Solid
General AI assistant
ChatGPT
68/100
ROLEGeneral-purpose assistant baseline and multimodal draft critic
Editorial fit
68
Source quality
55
Citation honesty
52
Privacy controls
40
Value for money
70
Speed
58
FREEFree
$8/MOGo
$20/MOPlus
SolidVerified 2026-06-09VP·METHOD

ChatGPT flagship-ready dossier: General-purpose assistant baseline and multimodal draft critic.

What is ChatGPT?

ChatGPT is a tool in the VerdictPal General AI assistant set: OpenAI’s general-purpose assistant for drafting, analysis, multimodal exploration, file work, coding, agents, and baseline AI comparison, with separate consumer, Business/Enterprise, and API trust boundaries. It scores 68 out of 100 on editorial fit.

Role: General-purpose assistant baseline and multimodal draft criticCategory: General AI assistantEditorial · hands-on
ChatGPT at a glance, with the date each field was checked.
FieldValue
Editorial fit68 out of 100
Dossier statusSolid
SetGeneral AI assistant
Role in a workflowGeneral-purpose assistant baseline and multimodal draft critic
PricingFree Free · Go $8/mo · Plus $20/mo · Pro $200/mo · Business Custom · Enterprise Custom
Training on your dataNot recorded as a plan-wide guarantee. Check the dossier's privacy notes and sources.
Last verified2026-06-09

Editor's note

ChatGPT is the default comparison point, which makes it easy to overrate. The tool should ask a sharper question: when does the broad assistant beat a specialized research, writing, coding, or citation tool — and when does it just make unsupported work sound finished?

Public facts · vendor & repo

List prices and pay-as-you-go entry points we can cite without running our own bench. Each tile links to a source when possible.

General AI assistantSetVerdictPal tool
synthesizedEvidenceQuality gate
6Sources checkedSources
2026-06-07Pricing checkedQuality gate

Limits & product surface

Non-price vendor claims — multipliers, caps, and API scope. Detailed matrices live in subscription and SDK sections below.

Primary surface
consumer-appTool identity
Modalities
text, image, audio, video, dataTool identity
Workflow roles
[object Object], [object Object], [object Object]VerdictPal editorial
Alternatives tracked
Claude, Gemini, Mistral Le Chat, Microsoft Copilot, OpenRouterVerdictPal comparison set

How it works · agent loop

The public positioning for this product — the loop we score against on VerdictPal.

Start with the wedge

General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against

Run the representative task

Run source-grounded answer, file summary, code fix, image read, and study-plan draft against ChatGPT, Claude, Gemini, Mistral, and DeepSeek.

Check the failure modes

Breadth can hide weak methodology

Compare before recommending

Compare against Claude, Gemini, Mistral Le Chat before shipping advice.

Who it fits, where it fails

Best for

  • General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against
  • Students who want one broad assistant for outlining, code help, file questions, images, audio, and study planning
  • Teams that can use Business/Enterprise controls rather than consumer chat for sensitive work
  • Workflows where connected services, GPTs/projects, Codex, and agents matter more than a narrow cited-search interface

Avoid if

  • You need API-only usage and are evaluating API economics, not the chat subscription
  • Institutional data rules forbid consumer cloud processing or model-improvement use
  • You require no-training and retention guarantees without a Business, Enterprise, Edu, Healthcare, Teachers, or API agreement
  • Your deliverable is source-grounded research and you will not independently open and verify citations

Strengths

  • General chat/reasoning
  • file uploads
  • multimodal text/image/audio/video/data
  • voice/video
  • Deep Research
  • ChatGPT Agent
  • Codex
  • Projects
  • GPTs
  • apps/connectors
  • Excel/Sheets extensions
  • Business admin controls
  • separate OpenAI API

Weaknesses

  • Breadth can hide weak methodology
  • Consumer content can train models unless controls are changed
  • Business admins may access workspace content
  • Subscription and API economics are separate
  • Usage caps/guardrails/flexible pricing can interrupt workflows
  • Citations and browsing need independent verification

How it could improve

  • Citation accuracy needs investment. A citation-verification pass against Semantic Scholar or CrossRef before presenting a source would catch the worst fabrication errors.
  • Source grounding scores 52/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Privacy posture scores 40/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.

Search modes · 4 lenses

One input box, many retrieval postures. Filter by tier.

01Free

Default

Default path for ChatGPT — verify limits on the live product.

02Pro

Deep work

Deep work path for ChatGPT — verify limits on the live product.

03Pro

Focused task

Focused task path for ChatGPT — verify limits on the live product.

04Free

Export & share

Export & share path for ChatGPT — verify limits on the live product.

Competitive lens

Where ChatGPT wins for cited research — and where a rival still belongs in the stack.

ChatGPT is stronger when general-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against; Claude may still win for narrower fit, procurement, or specialist depth.

ChatGPT

  • General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against
  • Students who want one broad assistant for outlining, code help, file questions, images, audio, and study planning

Claude

  • You need API-only usage and are evaluating API economics, not the chat subscription
  • Institutional data rules forbid consumer cloud processing or model-improvement use

Scores and evidence

Metric lab · 16 dimensions

Click a tile for the editorial note. Color follows score: coral, yellow, mint.

Shape

Avg 72 · 40–100

68Editorial fit

Weighted roll-up across 13 dimensions for General AI assistant; pending desk verification if rescored from public facts.

By group

Fit83
  • Editorial fit68
  • Wedge task fit82
  • Feature depth100
Cost69
  • Free-tier utility71
  • Cost-to-value70
  • Opportunity cost65
Trust69
  • Source grounding52
  • Privacy posture40
  • Failure transparency98
  • Evidence strength55
  • Transparency98
Workflow70
  • Integration reach80
  • Setup friction58
  • Reliability70
  • Competitive position75
  • Data portability65

Benchmark ledger

Public rows are vendor or third-party claims we logged with a date. Desk rows are reserved for VerdictPal self-run results.

MeasureResultSource
Run source-grounded answer, file summary, code fix, image read, and study-plan draft against ChatGPT, Claude, Gemini, Mistral, and DeepSeekcorrectness, verification behavior, citation discipline, edit quality, multimodal accuracy, latency, quota friction.PlannedVerdictPal benchmark plan
Inspect consumer data controls, Temporary Chat, Business workspace settings, and API termsdefault clarity, opt-out clarity, admin access, retention controls, export/delete path.PlannedVerdictPal benchmark plan
Price the same 20-task workload through ChatGPT subscription and API callsannual cost, task throughput, rate limits, governance, developer overhead.PlannedVerdictPal benchmark plan

Editorial evidence · 0 entries

No editorial benchmark yet. The tool stays at status Solid until evidence lands.

This tool is in the VerdictPal Citation Fidelity protocol — frozen question bank, dual-reviewer grading, results still pending. Citation Fidelity v0.1 →

Test scenarios · hands-on lab

How we exercised the product. Step through each run before you trust the scores.

Test run

Step 1 of 4

Start from the persona in best-for item 1.

Research log · desk notes

What we learned while building this dossier — not vendor copy.

Notion Tool pipeline (Card pipeline)

Notion research pass

ChatGPT is the baseline general assistant tool, but it now needs to be framed as a family of surfaces: consumer ChatGPT, Business/Enterprise workspaces, Codex/agent features, and the separate OpenAI API. Its strength is breadth — drafting, file work, multimodal analysis, coding, agents, and extensions — while its risk is that fluent output can make weak sourcing, privacy assumptions, or plan limits invisible.

VerdictPal git

Flagship-ready structure

Converted imported research into metrics, panes, scenarios, benchmark rows, comparison notes, and pricing deck.

What it costs, what it keeps

Pricing deck · checked 2026-06-09

Plus

$20

month

  • Imported from the current pricing summary.
  • Verify live regional checkout before publishing procurement advice.

Pro

$200

month

  • Highest usage limits, all models including o1 Pro
  • Verify live regional checkout before publishing procurement advice.

RAW · checked 2026-06-09: Consumer Free/Go/Plus/Pro tiers on chatgpt.com · Business standard seat: USD 25/user/mo monthly or USD 20/user/mo annual in many countries, 2-seat minimum · Business/Enterprise flexible pricing and Codex-only seat caveats · OpenAI API separate per-token pricing.

Privacy deep-dive · checked 2026-06-09

OpenAI collects account, content, upload, connected-service, and usage data for consumer services. Content may be used to improve/train models subject to controls. Temporary Chat is not in history or used to improve models. Business/API offerings are governed separately and business/API data is not used for training by default.

Training on your dataNot recorded as a plan-wide guarantee. Check the dossier's privacy notes and sources.
EU data residencyNot recorded as a plan-wide guarantee. Check the dossier's privacy notes and sources.
SOC 2 attestationNot recorded as a plan-wide guarantee. Check the dossier's privacy notes and sources.
Local-first by defaultNot recorded as a plan-wide guarantee. Check the dossier's privacy notes and sources.

Under the hood

Vendors named on the product about page — useful for procurement and privacy reviews.

general assistant baseline
ChatGPT
draft critique
Claude
multimodal exploration
Gemini

Alternatives and context

Workflow roles

How this tool fits into a composed research stack:

general assistant baselinedraft critiquemultimodal exploration

Deep panes · 6 lenses

Editorial lenses only. Subscription and API pricing live in their own sections above.

Sources and provenance

Quality gate · Solid

Evidence ready
Benchmark pending

Verdict history

  • ChatGPT is the baseline general assistant tool, but it now needs to be framed as a family of surfaces: consumer ChatGPT, Business/Enterprise workspaces, Codex/agent features, and the separate OpenAI API. Its strength is breadth — drafting, file work, multimodal analysis, coding, agents, and extensions — while its risk is that fluent output can make weak sourcing, privacy assumptions, or plan limits invisible.
  • Converted imported Notion research into a full flagship-ready dossier template with metrics, panes, pricing deck, scenarios, benchmark rows, and comparison slices.

Related dossiers

Browse the atlas

Questions this dossier answers

Every answer below is assembled from the dated fields on this page. Nothing is written separately for search.

Is ChatGPT worth using?

ChatGPT scores 68 out of 100 on editorial fit and carries Solid dossier status. Its job in a research workflow is: General-purpose assistant baseline and multimodal draft critic.

Who is ChatGPT best for?

ChatGPT earns its place when you need:

  • General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against
  • Students who want one broad assistant for outlining, code help, file questions, images, audio, and study planning
  • Teams that can use Business/Enterprise controls rather than consumer chat for sensitive work
  • Workflows where connected services, GPTs/projects, Codex, and agents matter more than a narrow cited-search interface

When should you not use ChatGPT?

Skip ChatGPT in these cases:

  • You need API-only usage and are evaluating API economics, not the chat subscription
  • Institutional data rules forbid consumer cloud processing or model-improvement use
  • You require no-training and retention guarantees without a Business, Enterprise, Edu, Healthcare, Teachers, or API agreement
  • Your deliverable is source-grounded research and you will not independently open and verify citations

What does ChatGPT cost?

Free Free · Go $8/mo · Plus $20/mo · Pro $200/mo · Business Custom · Enterprise Custom. Pricing last checked 2026-06-09.

Does ChatGPT train on your data?

Not recorded as a plan-wide guarantee. Check the dossier's privacy notes and sources.. Privacy terms last checked 2026-06-09.

What goes wrong with ChatGPT?

The dossier publishes 6 failure modes for ChatGPT, and they stay published whether or not the vendor likes them:

  • Breadth can hide weak methodology
  • Consumer content can train models unless controls are changed
  • Business admins may access workspace content
  • Subscription and API economics are separate
  • Usage caps/guardrails/flexible pricing can interrupt workflows
  • Citations and browsing need independent verification

What are the alternatives to ChatGPT?

The closest options to ChatGPT are Claude, Gemini, Mistral Le Chat, Microsoft Copilot, OpenRouter. Each one that has a tool in the atlas is linked from this dossier, with a head-to-head comparison.

The Pack · Editorial newsletter

New tools in your inbox. Free.

One short email when a tool ships or changes status. No tracking, no third-party analytics. Unsubscribe in one click.