← All cardsDOSSIER · General AI assistant · SOLID · VERIFIED 2026-06-09
Microsoft Copilot42SolidBenchmark pendingMicrosoft ecosystem assistant and Office-integrated AI layerVerified 2026-06-09

Dossier · General AI assistant

Microsoft Copilot

Microsoft ecosystem assistant and Office-integrated AI layer · last verified 2026-06-09

Solid
General AI assistant
Microsoft Copilot
42/100
ROLEMicrosoft ecosystem assistant and Office-integrated AI layer
Editorial fit
42
Source quality
58
Citation honesty
52
Privacy controls
40
Value for money
18
Speed
74
FREECopilot (free)
$20/MOCopilot Pro
CUSTOMper seat
SolidVerified 2026-06-09VP·METHOD

VERDICT HISTORY

  • Microsoft Copilot is a family-name card, not a single-product card. This pass uses Microsoft Learn pages for Microsoft 365 Copilot and Copilot Chat data protection plus Microsoft pricing search results. The key editorial job is to separate consumer Copilot, Microsoft 365 Copilot Chat, paid Microsoft 365 Copilot, and Azure/OpenAI developer paths.
  • Converted imported Notion research into a full flagship-ready dossier template with metrics, panes, pricing deck, scenarios, benchmark rows, and comparison slices.

EDITOR'S NOTE

Desk hands-on 2026-06: Post-pricing-update value is poor for most buyers — the card reflects near-giveaway cost-to-value until Microsoft clarifies what Pro still includes.

AT A GLANCE

Microsoft’s Copilot family across consumer Copilot, Microsoft 365 Copilot Chat, Microsoft 365 Copilot in Office apps, agents, Edge/Windows surfaces, and Azure/OpenAI-adjacent developer workflows.

Role: Microsoft ecosystem assistant and Office-integrated AI layerCategory: General AI assistantEditorial · hands-on

Microsoft Copilot flagship-ready dossier: Microsoft ecosystem assistant and Office-integrated AI layer.

PUBLIC FACTS · vendor & repo

List prices and pay-as-you-go entry points we can cite without running our own bench. Each tile links to a source when possible.

General AI assistantSetVerdictPal card
synthesizedEvidenceQuality gate
4Sources checkedCard sources
2026-06-07Pricing checkedQuality gate

Limits & product surface

Non-price vendor claims — multipliers, caps, and API scope. Detailed matrices live in subscription and SDK sections below.

Primary surface
consumer-appCard identity
Modalities
text, image, codeCard identity
Workflow roles
[object Object], [object Object], [object Object]VerdictPal editorial
Alternatives tracked
ChatGPT, Claude, Google Gemini, Perplexity, Microsoft 365VerdictPal comparison set

HOW IT WORKS · agent loop

The public positioning for this product — the loop we score against on VerdictPal.

Start with the wedge

Users already working inside Word, Excel, PowerPoint, Outlook, Teams, SharePoint, Edge, or Windows

Run the representative task

Run the same tasks in consumer Copilot, Copilot Chat with work account, and paid M365 Copilot.

Check the failure modes

Copilot products have different data boundaries

Compare before recommending

Compare against ChatGPT, Claude, Google Gemini before shipping advice.

METRIC LAB · 16 DIMENSIONS

Click a tile for the editorial note. Color follows score: coral, yellow, mint.

Shape

Avg 54 · 18–98

42Editorial fit

Weighted roll-up across 13 dimensions for General AI assistant; pending desk verification if rescored from public facts.

By group

Fit49
  • Editorial fit42
  • Wedge task fit48
  • Feature depth58
Cost34
  • Free-tier utility38
  • Cost-to-value18
  • Opportunity cost45
Trust65
  • Source grounding52
  • Privacy posture40
  • Failure transparency76
  • Evidence strength58
  • Transparency98
Workflow57
  • Integration reach62
  • Setup friction74
  • Reliability48
  • Competitive position40
  • Data portability59

SEARCH MODES · 4 lenses

One input box, many retrieval postures. Filter by tier.

01Free

Default

Default path for Microsoft Copilot — verify limits on the live product.

02Pro

Deep work

Deep work path for Microsoft Copilot — verify limits on the live product.

03Pro

Focused task

Focused task path for Microsoft Copilot — verify limits on the live product.

04Free

Export & share

Export & share path for Microsoft Copilot — verify limits on the live product.

BENCHMARK LEDGER

Public rows are vendor or third-party claims we logged with a date. Desk rows are reserved for VerdictPal self-run results.

MeasureResultSource
Run the same tasks in consumer Copilot, Copilot Chat with work account, and paid M365 Copilotdata boundary clarity, citation/source access, output quality, admin controls, and source-document verification effort.PlannedVerdictPal benchmark plan
Use a controlled SharePoint/Teams tenant with intentionally messy permissionssurfaced documents, permission leak risk, Purview visibility, admin remediation path.PlannedVerdictPal benchmark plan
Word rewrite, Excel analysis, PowerPoint draft, Outlook triage, Teams recap, and web-grounded answerPending benchmark row from Notion.PlannedVerdictPal benchmark plan

PRICING DECK · checked 2026-06-09

Microsoft 365 Copilot

Custom

month

  • Requires qualifying M365 plan
  • Unit: per seat

RAW · checked 2026-06-09: Microsoft 365 Copilot Chat is listed as available at no additional cost for eligible Microsoft Entra users with an eligible Microsoft 365 subscription; agents require Azure subscription or Copilot Studio capacity. M365 Copilot business pricing/bundles vary by tenant, plan, region, and licensing route; consumer Copilot/Pro are separate.

RESEARCH LOG · desk notes

What we learned while building this dossier — not vendor copy.

Notion Card pipeline

Notion research pass

Microsoft Copilot is a family-name card, not a single-product card. This pass uses Microsoft Learn pages for Microsoft 365 Copilot and Copilot Chat data protection plus Microsoft pricing search results. The key editorial job is to separate consumer Copilot, Microsoft 365 Copilot Chat, paid Microsoft 365 Copilot, and Azure/OpenAI developer paths.

VerdictPal git

Flagship-ready structure

Converted imported research into metrics, panes, scenarios, benchmark rows, comparison notes, and pricing deck.

COMPETITIVE LENS

Where Microsoft Copilot wins for cited research — and where a rival still belongs in the stack.

Microsoft Copilot is stronger when users already working inside word, excel, powerpoint, outlook, teams, sharepoint, edge, or windows; ChatGPT may still win for narrower fit, procurement, or specialist depth.

Microsoft Copilot

  • Users already working inside Word, Excel, PowerPoint, Outlook, Teams, SharePoint, Edge, or Windows
  • Organizations with Microsoft 365 governance, Purview, Entra ID, SharePoint permissions, and admin controls already in place

ChatGPT

  • You need one simple product with identical guarantees across personal and work accounts
  • Your institution has not approved Microsoft 365 Copilot for thesis, student-data, or regulated workflows

UNDER THE HOOD

Vendors named on the product about page — useful for procurement and privacy reviews.

Office-integrated drafting
Microsoft Copilot
web search answers
ChatGPT
Windows/Edge assistant
Claude

DEEP PANES · 6 LENSES

Editorial lenses only. Subscription and API pricing live in their own sections above.

TEST SCENARIOS · hands-on lab

How we exercised the product. Step through each run before you trust the scores.

Test run

Step 1 of 4

Start from the persona in best-for item 1.

BEST FOR

  • Users already working inside Word, Excel, PowerPoint, Outlook, Teams, SharePoint, Edge, or Windows
  • Organizations with Microsoft 365 governance, Purview, Entra ID, SharePoint permissions, and admin controls already in place
  • Work accounts that benefit from enterprise data protection, activity history, eDiscovery, retention, and compliance tooling
  • Students comparing consumer Copilot, Copilot Chat, ChatGPT, Gemini, and Claude for web answers/document drafting

AVOID IF

  • You need one simple product with identical guarantees across personal and work accounts
  • Your institution has not approved Microsoft 365 Copilot for thesis, student-data, or regulated workflows
  • Your tenant permissions are messy and users can access documents they should not see
  • You expect Office grounding to make answers correct without source-document verification

STRENGTHS

  • Consumer Copilot
  • M365 Copilot Chat with EDP
  • Copilot in Word/Excel/PowerPoint/Outlook/Teams
  • Graph grounding
  • activity history
  • Purview retention/eDiscovery
  • Bing web grounding
  • agents/connectors
  • Entra/Purview/sensitivity labels

WEAKNESSES

  • Copilot products have different data boundaries
  • Messy tenant permissions can expose content
  • Bing web grounding has different data handling
  • Agents/connectors add third-party terms and scopes
  • Generated content still needs verification

HOW IT COULD IMPROVE

  • Over-refusal is a calibration problem. The tool needs finer-grained policy categories so legitimate research tasks don't get swept up in blanket safety filters.
  • Extracted values that require manual verification suggest the extraction pipeline needs a confidence-checking layer, cross-reference extracted claims against the original passage before presenting them as fact.
  • Editorial fit scores 42/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Wedge task fit scores 48/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Free-tier utility scores 38/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Cost-to-value scores 18/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Opportunity cost scores 45/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Source grounding scores 52/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Privacy posture scores 40/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Reliability scores 48/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.
  • Competitive position scores 40/100, the weakest dimension on this dossier. Addressing this would meaningfully raise the editorial fit.

EDITORIAL EVIDENCE · 3 ENTRIES

TASK

Run the same tasks in consumer Copilot, Copilot Chat with work account, and paid M365 Copilot.

Pending.

data boundary clarity, citation/source access, output quality, admin controls, and source-document verification effort.

TASK

Use a controlled SharePoint/Teams tenant with intentionally messy permissions.

Pending.

surfaced documents, permission leak risk, Purview visibility, admin remediation path.

TASK

Word rewrite, Excel analysis, PowerPoint draft, Outlook triage, Teams recap, and web-grounded answer.

Pending.

Pending benchmark row from Notion.

This tool is in the VerdictPal Citation Fidelity protocol — frozen question bank, dual-reviewer grading, results still pending. Citation Fidelity v0.1 →

PRIVACY DEEP-DIVE · checked 2026-06-09

Microsoft Learn says M365 Copilot uses LLMs, Microsoft Graph content the user can access, and M365 apps. Prompts, responses, and Graph data are not used to train foundation LLMs; content remains in M365 service boundary; Azure OpenAI is used and does not cache customer content for M365 Copilot. Copilot Chat is web-grounded by default and logs prompts/responses to Exchange for auditing/eDiscovery.

Training on your dataNot stated
EU data residencyNot documented
SOC 2 attestationNot documented
Local-first by defaultCloud only

WORKFLOW ROLES

How this tool fits into a composed research stack:

Office-integrated draftingweb search answersWindows/Edge assistant

QUALITY GATE · SOLID

Evidence ready
Benchmark pending
  • desk hands-on 2026-06-15: headline and metrics adjusted after desk trial.
  • score-revision 2026-06-09: differentiated from public facts; pending desk verification
  • Flagship-quality promotion 2026-06-04: source-backed dossier promoted; benchmark and public recommendation remain locked until evidence packet is complete.
  • Next action: Run ecosystem split, permission review, and Office workflow benchmarks.
  • Has API: true
  • Public recommendation: false
  • evidence-override: Notion source-of-truth sync; evidence remains source-backed until desk benchmark or hands-on pass.
  • Flagship-ready structure generated from Notion deep research and imported git fields on 2026-06-09.

Related dossiers

Browse the atlas

The Pack · Editorial newsletter

New cards in your inbox. Free.

One short email when a card ships or changes status. No tracking, no third-party analytics. Unsubscribe in one click.