← All cardsDOSSIER · General AI assistant · SOLID · VERIFIED 2026-06-09
ChatGPT68SolidBenchmark pendingGeneral-purpose assistant baseline and multimodal draft criticVerified 2026-06-09

Dossier · General AI assistant

ChatGPT

General-purpose assistant baseline and multimodal draft critic · last verified 2026-06-09

Solid
General AI assistant
ChatGPT
68/100
ROLEGeneral-purpose assistant baseline and multimodal draft critic
Editorial fit
68
Source quality
55
Citation honesty
52
Privacy controls
40
Value for money
70
Speed
58
FREEFree
$8/MOGo
$20/MOPlus
SolidVerified 2026-06-09VP·METHOD

VERDICT HISTORY

  • ChatGPT is the baseline general assistant card, but it now needs to be framed as a family of surfaces: consumer ChatGPT, Business/Enterprise workspaces, Codex/agent features, and the separate OpenAI API. Its strength is breadth — drafting, file work, multimodal analysis, coding, agents, and extensions — while its risk is that fluent output can make weak sourcing, privacy assumptions, or plan limits invisible.
  • Converted imported Notion research into a full flagship-ready dossier template with metrics, panes, pricing deck, scenarios, benchmark rows, and comparison slices.

EDITOR'S NOTE

ChatGPT is the default comparison point, which makes it easy to overrate. The card should ask a sharper question: when does the broad assistant beat a specialized research, writing, coding, or citation tool — and when does it just make unsupported work sound finished?

AT A GLANCE

OpenAI’s general-purpose assistant for drafting, analysis, multimodal exploration, file work, coding, agents, and baseline AI comparison, with separate consumer, Business/Enterprise, and API trust boundaries.

Role: General-purpose assistant baseline and multimodal draft criticCategory: General AI assistantEditorial · hands-on

ChatGPT flagship-ready dossier: General-purpose assistant baseline and multimodal draft critic.

PUBLIC FACTS · vendor & repo

List prices and pay-as-you-go entry points we can cite without running our own bench. Each tile links to a source when possible.

General AI assistantSetVerdictPal card
synthesizedEvidenceQuality gate
6Sources checkedCard sources
2026-06-07Pricing checkedQuality gate

Limits & product surface

Non-price vendor claims — multipliers, caps, and API scope. Detailed matrices live in subscription and SDK sections below.

Primary surface
consumer-appCard identity
Modalities
text, image, audio, video, dataCard identity
Workflow roles
[object Object], [object Object], [object Object]VerdictPal editorial
Alternatives tracked
Claude, Gemini, Mistral Le Chat, Microsoft Copilot, OpenRouterVerdictPal comparison set

HOW IT WORKS · agent loop

The public positioning for this product — the loop we score against on VerdictPal.

Start with the wedge

General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against

Run the representative task

Run source-grounded answer, file summary, code fix, image read, and study-plan draft against ChatGPT, Claude, Gemini, Mistral, and DeepSeek.

Check the failure modes

Breadth can hide weak methodology

Compare before recommending

Compare against Claude, Gemini, Mistral Le Chat before shipping advice.

METRIC LAB · 16 DIMENSIONS

Click a tile for the editorial note. Color follows score: coral, yellow, mint.

Shape

Avg 72 · 40–100

68Editorial fit

Weighted roll-up across 13 dimensions for General AI assistant; pending desk verification if rescored from public facts.

By group

Fit83
  • Editorial fit68
  • Wedge task fit82
  • Feature depth100
Cost69
  • Free-tier utility71
  • Cost-to-value70
  • Opportunity cost65
Trust69
  • Source grounding52
  • Privacy posture40
  • Failure transparency98
  • Evidence strength55
  • Transparency98
Workflow70
  • Integration reach80
  • Setup friction58
  • Reliability70
  • Competitive position75
  • Data portability65

SEARCH MODES · 4 lenses

One input box, many retrieval postures. Filter by tier.

01Free

Default

Default path for ChatGPT — verify limits on the live product.

02Pro

Deep work

Deep work path for ChatGPT — verify limits on the live product.

03Pro

Focused task

Focused task path for ChatGPT — verify limits on the live product.

04Free

Export & share

Export & share path for ChatGPT — verify limits on the live product.

BENCHMARK LEDGER

Public rows are vendor or third-party claims we logged with a date. Desk rows are reserved for VerdictPal self-run results.

MeasureResultSource
Run source-grounded answer, file summary, code fix, image read, and study-plan draft against ChatGPT, Claude, Gemini, Mistral, and DeepSeekcorrectness, verification behavior, citation discipline, edit quality, multimodal accuracy, latency, quota friction.PlannedVerdictPal benchmark plan
Inspect consumer data controls, Temporary Chat, Business workspace settings, and API termsdefault clarity, opt-out clarity, admin access, retention controls, export/delete path.PlannedVerdictPal benchmark plan
Price the same 20-task workload through ChatGPT subscription and API callsannual cost, task throughput, rate limits, governance, developer overhead.PlannedVerdictPal benchmark plan

PRICING DECK · checked 2026-06-09

Plus

$20

month

  • Imported from the current pricing summary.
  • Verify live regional checkout before publishing procurement advice.

Pro

$200

month

  • Highest usage limits, all models including o1 Pro
  • Verify live regional checkout before publishing procurement advice.

RAW · checked 2026-06-09: Consumer Free/Go/Plus/Pro tiers on chatgpt.com · Business standard seat: USD 25/user/mo monthly or USD 20/user/mo annual in many countries, 2-seat minimum · Business/Enterprise flexible pricing and Codex-only seat caveats · OpenAI API separate per-token pricing.

RESEARCH LOG · desk notes

What we learned while building this dossier — not vendor copy.

Notion Card pipeline

Notion research pass

ChatGPT is the baseline general assistant card, but it now needs to be framed as a family of surfaces: consumer ChatGPT, Business/Enterprise workspaces, Codex/agent features, and the separate OpenAI API. Its strength is breadth — drafting, file work, multimodal analysis, coding, agents, and extensions — while its risk is that fluent output can make weak sourcing, privacy assumptions, or plan limits invisible.

VerdictPal git

Flagship-ready structure

Converted imported research into metrics, panes, scenarios, benchmark rows, comparison notes, and pricing deck.

COMPETITIVE LENS

Where ChatGPT wins for cited research — and where a rival still belongs in the stack.

ChatGPT is stronger when general-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against; Claude may still win for narrower fit, procurement, or specialist depth.

ChatGPT

  • General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against
  • Students who want one broad assistant for outlining, code help, file questions, images, audio, and study planning

Claude

  • You need API-only usage and are evaluating API economics, not the chat subscription
  • Institutional data rules forbid consumer cloud processing or model-improvement use

UNDER THE HOOD

Vendors named on the product about page — useful for procurement and privacy reviews.

general assistant baseline
ChatGPT
draft critique
Claude
multimodal exploration
Gemini

DEEP PANES · 6 LENSES

Editorial lenses only. Subscription and API pricing live in their own sections above.

TEST SCENARIOS · hands-on lab

How we exercised the product. Step through each run before you trust the scores.

Test run

Step 1 of 4

Start from the persona in best-for item 1.

BEST FOR

  • General-purpose drafting, critique, summarization, and multimodal exploration when you need the baseline assistant to compare against
  • Students who want one broad assistant for outlining, code help, file questions, images, audio, and study planning
  • Teams that can use Business/Enterprise controls rather than consumer chat for sensitive work
  • Workflows where connected services, GPTs/projects, Codex, and agents matter more than a narrow cited-search interface

AVOID IF

  • You need API-only usage and are evaluating API economics, not the chat subscription
  • Institutional data rules forbid consumer cloud processing or model-improvement use
  • You require no-training and retention guarantees without a Business, Enterprise, Edu, Healthcare, Teachers, or API agreement
  • Your deliverable is source-grounded research and you will not independently open and verify citations

STRENGTHS

  • General chat/reasoning
  • file uploads
  • multimodal text/image/audio/video/data
  • voice/video
  • Deep Research
  • ChatGPT Agent
  • Codex
  • Projects
  • GPTs
  • apps/connectors
  • Excel/Sheets extensions
  • Business admin controls
  • separate OpenAI API

WEAKNESSES

  • Breadth can hide weak methodology
  • Consumer content can train models unless controls are changed
  • Business admins may access workspace content
  • Subscription and API economics are separate
  • Usage caps/guardrails/flexible pricing can interrupt workflows
  • Citations and browsing need independent verification

HOW IT COULD IMPROVE

  • Citation accuracy needs investment. A citation-verification pass against Semantic Scholar or CrossRef before presenting a source would catch the worst fabrication errors.
  • Source grounding scores 52/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Privacy posture scores 40/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.

EDITORIAL EVIDENCE · 3 ENTRIES

TASK

Run source-grounded answer, file summary, code fix, image read, and study-plan draft against ChatGPT, Claude, Gemini, Mistral, and DeepSeek.

Pending.

correctness, verification behavior, citation discipline, edit quality, multimodal accuracy, latency, quota friction.

TASK

Inspect consumer data controls, Temporary Chat, Business workspace settings, and API terms.

Pending.

default clarity, opt-out clarity, admin access, retention controls, export/delete path.

TASK

Price the same 20-task workload through ChatGPT subscription and API calls.

Pending.

annual cost, task throughput, rate limits, governance, developer overhead.

This tool is in the VerdictPal Citation Fidelity protocol — frozen question bank, dual-reviewer grading, results still pending. Citation Fidelity v0.1 →

PRIVACY DEEP-DIVE · checked 2026-06-09

OpenAI collects account, content, upload, connected-service, and usage data for consumer services. Content may be used to improve/train models subject to controls. Temporary Chat is not in history or used to improve models. Business/API offerings are governed separately and business/API data is not used for training by default.

Training on your dataNot stated
EU data residencyNot documented
SOC 2 attestationNot documented
Local-first by defaultCloud only

APPEARS IN

Composed workflows on VerdictPal that reference this card, not vendor marketing.

Stacks

  • Grant deadline stackPhD students, postdocs, and PIs racing a real submission deadline with limited time for tool drama.

WORKFLOW ROLES

How this tool fits into a composed research stack:

general assistant baselinedraft critiquemultimodal exploration

QUALITY GATE · SOLID

Evidence ready
Benchmark pending
  • desk hands-on 2026-07-05: personally tried; evidence upgraded from synthesized.
  • score-revision 2026-06-09: differentiated from public facts; pending desk verification
  • Flagship-quality promotion 2026-06-04: source-backed dossier promoted; benchmark and public recommendation remain locked until evidence packet is complete.
  • Next action: Run baseline assistant benchmark against Claude, Gemini, Mistral, and DeepSeek; include privacy/settings and API-versus-chat economics audits.
  • Has API: true
  • Public recommendation: false
  • Flagship-ready structure generated from Notion deep research and imported git fields on 2026-06-09.

Related dossiers

Browse the atlas

The Pack · Editorial newsletter

New cards in your inbox. Free.

One short email when a card ships or changes status. No tracking, no third-party analytics. Unsubscribe in one click.