Skip to content

Rank #9 of 9 in AI Assistants

Microsoft Copilot

Built-in AI assistant

Microsoft · commercial

no public signals

Microsoft ships more than one product — each judged line competes in its own arena on the same stories as everyone else.

LineArenaRankPA Score
Microsoft Copilotthis pageAI Assistants#9/94/100
WindowsDesktop OS#3/59/100
Microsoft TeamsTeam Chat#5/521/100
AutoGenAgent Frameworks & SDKs#9/926/100

Not yet judged (1 — no arena where they compete): Visual Studio Code

Verified integrations

No integration evidence found in our corpus for this product yet — that means none was found, never that it doesn’t integrate.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

7.4/100

Agents tasks — stories about agents tasks in this arenaAgents tasksevidence →

Stories about agents tasks in this arena

27.0/100

Apps devices — stories about apps devices in this arenaApps devicesevidence →

Stories about apps devices in this arena

12.0/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

4.5/100

Connectors apps — stories about connectors apps in this arenaConnectors appsevidence →

Stories about connectors apps in this arena

57.6/100

Files analysis — stories about files analysis in this arenaFiles analysisevidence →

Stories about files analysis in this arena

30.0/100

Memory context — stories about memory context in this arenaMemory contextevidence →

Stories about memory context in this arena

23.0/100

Multimodal — stories about multimodal in this arenaMultimodalevidence →

Stories about multimodal in this arena

30.0/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

0.0/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

23.3/100

Research answers — stories about research answers in this arenaResearch answersevidence →

Stories about research answers in this arena

50.0/100

Trust controls — stories about trust controls in this arenaTrust controlsevidence →

Stories about trust controls in this arena

18.0/100

Story verdicts — every judged story with its evidenceStory verdicts

?

Sorted by importance (agentic first) (high → low) · 52/52 stories · click a row’s chevron for the rationale and evidence

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3partial6/10X

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3n/a0/10

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none0/10

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none0/10

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2disputed5/10D

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2disputed5/10D

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial4/10C

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1n/auntestednone yet

Connect my cloud drive, email, and calendar so the assistant can search and use them in answers C

Connectors

knowledge-workerConnectors apps — stories about connectors apps in this arenaConnectors apps3full8/10X

Control whether my conversations are used to train models C

Data controls

knowledge-workerTrust controls — stories about trust controls in this arenaTrust controls3partial6/10X

Have the assistant operate a web browser on my behalf to research and complete tasks on websites C

Agent mode

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks3partial6/10C

Have the assistant remember relevant context from previous chats and apply it in new conversations C

Memory

power-userMemory context — stories about memory context in this arenaMemory context3partial6/10C

Have the assistant write and run code on my data to produce charts, computed answers, and downloadable files C

Analysis

power-userFiles analysis — stories about files analysis in this arenaFiles analysis3partial6/10C

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3partial6/10X

Upload documents, spreadsheets, and PDFs and get accurate analysis of their contents C

Files

knowledge-workerFiles analysis — stories about files analysis in this arenaFiles analysis3partial6/10X

Delegate a multi-step task that the assistant works on autonomously in the background and returns for my review C

Agent mode

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks3partial5/10C

Launch a deep research run that autonomously searches many sources and returns a cited report C

Research

knowledge-workerResearch answers — stories about research answers in this arenaResearch answers3partial5/10C

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3none0/10

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3noneuntestednone yet

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3noneuntestednone yet

Get answers grounded in current web results with citations back to the sources C

Research

knowledge-workerResearch answers — stories about research answers in this arenaResearch answers2full8/10C

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2partial6/10X

Have a natural, real-time voice conversation with the assistant C

Voice

knowledge-workerMultimodal — stories about multimodal in this arenaMultimodal2full6/10C

Let the assistant see and operate applications on my computer to complete work C

Agent mode

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks2partial6/10X

Use full-featured official mobile apps for iOS and Android C

Apps

knowledge-workerApps devices — stories about apps devices in this arenaApps devices2partial6/10C

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2disputed5/10D

Share screenshots and photos and have the assistant accurately interpret what is in them C

Images

knowledge-workerMultimodal — stories about multimodal in this arenaMultimodal2partial5/10C

Browse a directory of third-party apps and connectors and add them to the assistant C

Connectors

power-userConnectors apps — stories about connectors apps in this arenaConnectors apps2partial4/10C

Have the assistant create and iteratively edit documents, presentations, and other files I can export G

Artifacts

knowledge-workerFiles analysis — stories about files analysis in this arenaFiles analysis2disputed4/10D

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2partial3/10X

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2none0/10

Generate and edit images from natural-language prompts C

Images

knowledge-workerMultimodal — stories about multimodal in this arenaMultimodal2none0/10

Organize related chats and files into a project or space that shares context and instructions C

Projects

knowledge-workerMemory context — stories about memory context in this arenaMemory context2none0/10

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2none0/10

Schedule recurring or one-off tasks that run automatically and come back to me with results C

Tasks

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks2none0/10

Use an official desktop app with OS-level shortcuts and access to what is on my screen C

Apps

power-userApps devices — stories about apps devices in this arenaApps devices2none0/10

Build and share custom assistants with their own instructions and knowledge C

Custom bots

power-userApps devices — stories about apps devices in this arenaApps devices2noneuntestednone yet

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Manage members, permissions, and data policies for my organization's workspace C

Admin

team-adminTrust controls — stories about trust controls in this arenaTrust controls2noneuntestednone yet

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2noneuntestednone yet

Set persistent custom instructions and preferences that shape every response C

Memory

power-userMemory context — stories about memory context in this arenaMemory context1partial5/10C

Export my complete chat history and account data G

Data controls

knowledge-workerTrust controls — stories about trust controls in this arenaTrust controls1none0/10

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1noneuntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 41 stories with headroom

What would move Microsoft Copilot’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Agenticness — how well agents can access and operate the productPlug MCP servers into this product so it can use their tools

    nonemoves agent-readyimpact 45

    The evidence pack describes Copilot connecting to first-party consumer services (OneDrive, Outlook.com, Google Drive/Gmail) via 'Copilot connectors,' but there is no mention of MCP (Model Context Protocol) server support or an ability for users to plug in arbitrary MCP tool servers.

  2. Agenticness — how well agents can access and operate the productDrive the product through a documented public API

    nonemoves agent-readyimpact 45

    Missing: any public API documentation, SDK, or endpoint spec that lets developers programmatically invoke Copilot's capabilities.

  3. Automation depth — how much of the product can run unattendedDefine rules that trigger actions automatically on events

    nonemoves PA Scoreimpact 30

    The evidence describes Copilot responding to prompts, browsing, summarizing, and connecting to services, but there is no mention of user-defined rules or triggers that fire actions automatically on events (e.g., 'when email arrives, do X' style automation).

  4. Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave

    nonemoves PA Scoreimpact 30

    No evidence of a data export/portability feature or open-format export mechanism; docs only mention deleting conversation history, not exporting it, and probes for machine-readable docs/APIs 404.

  5. Openness — open source, data portability, and self-hosting storiesSelf-host the core product

    nonemoves PA Scoreimpact 30

    Microsoft Copilot is a cloud-hosted SaaS product with no evidence of any self-hosted or on-premises deployment option; all documentation describes it as a hosted service (Copilot.com, Microsoft 365 cloud integration, Edge browser actions) with no self-hosting mechanism mentioned.

  6. Agenticness — how well agents can access and operate the productPoint an agent at llms.txt or agent-oriented docs

    nonemoves agent-readyimpact 30

    Probes show no llms.txt, no markdown docs endpoint, and no OpenAPI spec for Microsoft Copilot's support site, and no evidence anywhere in the pack of agent-oriented documentation formats being offered.

  7. Agenticness — how well agents can access and operate the productUse an official CLI

    nonemoves agent-readyimpact 30

    No evidence of an official CLI for Microsoft Copilot; all evidence covers chat, browser, and Office app integrations, with probes confirming no machine-readable API/CLI artifacts.

  8. Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agent

    nonemoves agent-readyimpact 30

    Missing: any mention of API keys, scoped tokens, permission scoping, or credential management for agent access.

Showing the top 8 of 41 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map3 surfaces · 24 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

En us docs23 stories

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

2 of 16 testable claims verified · 3 contradictedintegrity 0/100

25 distinct capability claims found in Microsoft Copilot’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

2

Verified

11

Unverified

3

Contradicted

7

Undersold

Verified (6)
Unverified (15)
Contradicted (5)
Undersold (7)
Claims outside our story set (1)

Real capability claims found in Microsoft Copilot’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Users can choose between different leading AI models for different kinds of work

    source ↗
Suggest a story for these →

Business model

free-tiersubscription-flat

Free to use with a Microsoft account; Copilot Pro is a flat monthly subscription that adds priority model access and Copilot in Microsoft 365 apps.

pricing ↗

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Scoretracked since Sep 14 '26 — no movement recorded yet
Agent-readytracked since Sep 14 '26 — no movement recorded yet

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data