What AI research tools actually cost, and what they do with your data
The short version
Across 52 published tool dossiers and 175 models, three findings hold. Paid entry plans cluster tightly around $11 a month while frontier model input prices spread 1153.8×× from cheapest to dearest. Only 4 of 52 dossiers record a default-off stance on training with your content. And of 29 tracked benchmarks, 7 no longer separate models well enough to be read as live evidence.
Free to cite with a link. The desk would rather be quoted accurately than quoted exclusively.
What does an AI research tool cost?
Across 52 published dossiers, 42 carry a usable free tier (81 percent). 30 publish a structured monthly price; the median paid entry plan is $11 a month, running from $1 at the low end to $49 for Elicit.
| Field | Value |
|---|---|
| Published tool dossiers | 52 |
| Carry a free tier | 42 / 52 (81%) |
| Publish a structured monthly price | 30 / 52 |
| Median paid entry plan | $11 / mo |
| Cheapest paid entry plan | $1 / mo |
| Dearest paid entry plan | $49 / mo — Elicit |
How many AI tools say what they do with your data?
4 of 52 dossiers record training on your content as off by default. 6 document an opt-out you have to go and find. On the remaining 42, the terms the desk read did not state a position at all — and that silence is the finding, not a gap in the research.
| Field | Value |
|---|---|
| Training off by default, documented | 4 / 52 |
| Opt-out available, reader must find it | 6 / 52 |
| No stance stated in the terms read | 42 / 52 |
| EU data residency documented for everyone | 1 / 52 |
| EU residency on enterprise plans only | 20 / 52 |
| SOC 2 documented | 6 / 52 |
| Local-first by default | 1 / 52 |
These rows describe what each dossier records after the desk read the vendor's published terms. They are not a legal audit, and a blank is a blank — not an accusation.
How wide is the price gap between frontier models?
Across 175 models, 137 publish a per-token price. The median is $1.40 per million input tokens and $7.50 per million output tokens. The gap between the cheapest and dearest published input price is 1153.8× — $0.07 against $75 for Claude Mythos Preview.
| Field | Value |
|---|---|
| Published models | 175 |
| Publish a per-token input price | 137 / 175 |
| Median price per 1M input tokens | $1.40 |
| Median price per 1M output tokens | $7.50 |
| Cheapest published input price | $0.07 |
| Dearest published input price | $75 — Claude Mythos Preview |
| Input price spread, cheapest to dearest | 1153.8× |
| Largest published context window | 10M — Llama 4 Scout |
How many AI benchmarks still separate models?
The atlas tracks 29 published benchmarks against 1789 sourced score rows covering 175 models. 22 are active. 2 are saturated and 5 are marked defunct, because a wall of near-perfect scores tells a reader nothing and leaving it up as live evidence is worse than removing it.
| Field | Value |
|---|---|
| Published benchmark pages | 29 |
| Sourced score rows in the ledger | 1789 |
| Models with at least one sourced row | 175 |
Saturation
| Field | Value |
|---|---|
| Active | 22 |
| Defunct | 5 |
| Saturated | 2 |
Contamination risk
| Field | Value |
|---|---|
| Low contamination | 13 |
| Medium contamination | 10 |
| Unknown contamination | 3 |
| High contamination | 3 |
Score rows come from Artificial Analysis and the Vals Index, and every row on a benchmark page shows its source and the date it was checked. The desk publishes no first-party benchmark scores until grading and two-reviewer sign-off are complete.
How fresh are these numbers?
52 of 52 dossiers had pricing re-checked within the last 90 days. Checks currently run from 2026-06-07 to 2026-06-15. Freshness is the whole product here: a pricing page with no timestamp is a guess wearing a suit.
| Field | Value |
|---|---|
| Pricing re-checked within 90 days | 52 / 52 |
| Pricing re-checked within 180 days | 52 / 52 |
| Oldest pricing check on any tool | 2026-06-07 |
| Newest pricing check on any tool | 2026-06-15 |
Evidence tier
| Field | Value |
|---|---|
| Editorial · hands-on | 41 |
| Synthesized | 11 |
Dossier status
| Field | Value |
|---|---|
| Solid | 43 |
| Flagship | 9 |
How these numbers are produced
Every figure is computed at render time from the published tools in this atlas. Draft tools are excluded, because a draft is a stub and counting stubs would inflate every number on the page.
- Drafts excluded. Counts cover tools the desk has actually finished.
- Privacy rows report what a dossier records after the desk read the vendor's published terms — not a legal audit.
- Benchmark rows are third-party, each carrying a source label and a check date on its benchmark page.
- The edition label tracks the newest pricing check in the atlas, not the day you loaded the page. A page that renumbers itself daily without new evidence is false freshness.
- The same figures are available as JSON for anyone who wants to check the arithmetic.
Cite this report
Quote any figure with a link back to this page. If a number here contradicts something you have measured, the desk wants to hear it — corrections get published.
Suggested citation
VerdictPal editorial desk. The Atlas Report, edition 2026-06. Retrieved from https://verdictpal.com/report