Skip to content

Rank #1 of 9 in AI Assistants

ChatGPT logo

OpenAI · commercial

no public signals

OpenAI ships more than one product — each judged line competes in its own arena on the same stories as everyone else.

LineArenaRankPA Score
ChatGPTthis pageAI Assistants#1/938/100
CodexAI Coding Agents#6/1333/100
Agents SDKAgent Frameworks & SDKs#2/937/100
PluginsAgent Skills & Extensions#4/519/100

Not yet judged (6 — no arena where they compete): ChatGPT Work · Image generation (GPT-Image-2.5) · Codex Security · ChatKit · Agents API · API Platform

Verified integrations

Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

49.1/100

Agents tasks — stories about agents tasks in this arenaAgents tasksevidence →

Stories about agents tasks in this arena

82.0/100

Apps devices — stories about apps devices in this arenaApps devicesevidence →

Stories about apps devices in this arena

45.3/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

58.3/100

Connectors apps — stories about connectors apps in this arenaConnectors appsevidence →

Stories about connectors apps in this arena

49.6/100

Files analysis — stories about files analysis in this arenaFiles analysisevidence →

Stories about files analysis in this arena

78.8/100

Memory context — stories about memory context in this arenaMemory contextevidence →

Stories about memory context in this arena

80.0/100

Multimodal — stories about multimodal in this arenaMultimodalevidence →

Stories about multimodal in this arena

80.0/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

7.2/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

4.0/100

Research answers — stories about research answers in this arenaResearch answersevidence →

Stories about research answers in this arena

53.6/100

Trust controls — stories about trust controls in this arenaTrust controlsevidence →

Stories about trust controls in this arena

8.0/100

Story verdicts — every judged story with its evidenceStory verdicts

?

Sorted by importance (agentic first) (high → low) · 52/52 stories · click a row’s chevron for the rationale and evidence

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full9/10C

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full8/10C

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full8/10C

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3partial4/10T

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full9/10X

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10C

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10C

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full7/10C

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full7/10T

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10T

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial4/10C

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial4/10C

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial3/10T

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1partial4/10C

Delegate a multi-step task that the assistant works on autonomously in the background and returns for my review C

Agent mode

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks3full8/10C

Have the assistant operate a web browser on my behalf to research and complete tasks on websites C

Agent mode

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks3full8/10C

Have the assistant remember relevant context from previous chats and apply it in new conversations C

Memory

power-userMemory context — stories about memory context in this arenaMemory context3full8/10C

Have the assistant write and run code on my data to produce charts, computed answers, and downloadable files C

Analysis

power-userFiles analysis — stories about files analysis in this arenaFiles analysis3full8/10X

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3full7/10C

Upload documents, spreadsheets, and PDFs and get accurate analysis of their contents C

Files

knowledge-workerFiles analysis — stories about files analysis in this arenaFiles analysis3full7/10C

Connect my cloud drive, email, and calendar so the assistant can search and use them in answers C

Connectors

knowledge-workerConnectors apps — stories about connectors apps in this arenaConnectors apps3partial6/10C

Launch a deep research run that autonomously searches many sources and returns a cited report C

Research

knowledge-workerResearch answers — stories about research answers in this arenaResearch answers3partial6/10C

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3none0/10

Control whether my conversations are used to train models C

Data controls

knowledge-workerTrust controls — stories about trust controls in this arenaTrust controls3noneuntestednone yet

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3noneuntestednone yet

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3n/auntestednone yet

Generate and edit images from natural-language prompts C

Images

knowledge-workerMultimodal — stories about multimodal in this arenaMultimodal2full9/10C

Have the assistant create and iteratively edit documents, presentations, and other files I can export G

Artifacts

knowledge-workerFiles analysis — stories about files analysis in this arenaFiles analysis2full9/10C

Let the assistant see and operate applications on my computer to complete work C

Agent mode

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks2full9/10C

Get answers grounded in current web results with citations back to the sources C

Research

knowledge-workerResearch answers — stories about research answers in this arenaResearch answers2full8/10X

Have a natural, real-time voice conversation with the assistant C

Voice

knowledge-workerMultimodal — stories about multimodal in this arenaMultimodal2full8/10C

Organize related chats and files into a project or space that shares context and instructions C

Projects

knowledge-workerMemory context — stories about memory context in this arenaMemory context2full8/10C

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2full8/10C

Schedule recurring or one-off tasks that run automatically and come back to me with results C

Tasks

power-userAgents tasks — stories about agents tasks in this arenaAgents tasks2full8/10C

Browse a directory of third-party apps and connectors and add them to the assistant C

Connectors

power-userConnectors apps — stories about connectors apps in this arenaConnectors apps2full7/10C

Share screenshots and photos and have the assistant accurately interpret what is in them C

Images

knowledge-workerMultimodal — stories about multimodal in this arenaMultimodal2full7/10C

Use full-featured official mobile apps for iOS and Android C

Apps

knowledge-workerApps devices — stories about apps devices in this arenaApps devices2full7/10C

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2partial6/10C

Use an official desktop app with OS-level shortcuts and access to what is on my screen C

Apps

power-userApps devices — stories about apps devices in this arenaApps devices2partial6/10C

Build and share custom assistants with their own instructions and knowledge C

Custom bots

power-userApps devices — stories about apps devices in this arenaApps devices2partial5/10C

Manage members, permissions, and data policies for my organization's workspace C

Admin

team-adminTrust controls — stories about trust controls in this arenaTrust controls2partial4/10C

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2partial3/10C

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial3/10T

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2n/auntestednone yet

Set persistent custom instructions and preferences that shape every response C

Memory

power-userMemory context — stories about memory context in this arenaMemory context1full8/10C

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1partial4/10C

Export my complete chat history and account data G

Data controls

knowledge-workerTrust controls — stories about trust controls in this arenaTrust controls1none0/10

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 24 stories with headroom

What would move ChatGPT’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave

    nonemoves PA Scoreimpact 30

    No evidence in the pack of a data export feature producing open/portable formats, nor any mention of account data export or deletion workflow.

  2. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    Missing: any reference to training-data opt-out settings, business/API data-usage policies, or privacy dashboard controls.

  3. Trust controls — stories about trust controls in this arenaControl whether my conversations are used to train models

    nonemoves PA Scoreimpact 30

    The evidence pack covers ChatGPT's agentic/feature capabilities (Codex, Work, MCP, Computer Use, etc.) but contains no documentation or mention of data controls, training opt-out settings, or 'Improve the model for everyone' toggles that let a user control whether their conversations are used for model training.

  4. Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples

    nonemoves API qualityimpact 30

    The evidence pack contains no documentation of an interactive API reference or runnable code examples; a direct probe for OpenAPI/swagger specs on the docs site returned 404 across all candidate paths, and no other citation mentions an API reference sandbox or runnable snippets.

  5. Agenticness — how well agents can access and operate the productDownload a machine-readable API spec (OpenAPI or equivalent)

    nonemoves API qualityimpact 30

    The probe explicitly checked for an OpenAPI/machine-readable spec at all standard locations and found only 404s, with no other evidence pack item showing a downloadable API spec for ChatGPT itself.

  6. Agenticness — how well agents can access and operate the productRely on versioned APIs with a documented deprecation policy

    nonemoves API qualityimpact 30

    No evidence of a versioned API or a documented deprecation policy is present; the OpenAPI probe explicitly returned 404s and none of the docs mention API versioning or deprecation practices.

  7. Agenticness — how well agents can access and operate the productDrive the product through a documented public API

    partialq4/10moves agent-readyimpact 27

    Missing: an explicit REST/GraphQL API reference for ChatGPT product actions, OpenAPI/swagger spec, and independent corroboration that non-Codex ChatGPT features are API-drivable.

  8. Agenticness — how well agents can access and operate the productBuild against official SDKs

    partialq3/10moves agent-readyimpact 21

    Missing: dedicated SDK documentation/reference pages, code examples, language support details, independent developer reports of building with the SDK.

Showing the top 8 of 24 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map6 surfaces · 41 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

docs36 stories

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

2 of 22 testable claims verified · 0 contradictedintegrity 9/100

55 distinct capability claims found in ChatGPT’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

2

Verified

20

Unverified

0

Contradicted

19

Undersold

Verified (4)
Unverified (50)
Undersold (19)
Claims outside our story set (2)

Real capability claims found in ChatGPT’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • IDE extension can create and edit files directly in your workspace

    source ↗
  • Generate interactive visualizations and build/share websites and apps via Sites

    source ↗
Suggest a story for these →

Business model

free-tiersubscription-flatsubscription-per-seatenterprise-custom

Free tier plus flat-rate Plus and Pro subscriptions; Business/Team is per-seat and Enterprise is custom, with ChatGPT Work usage drawing on shared workspace credits.

pricing ↗

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score34 (Sep 14 '26)38 (Sep 15 '26)
Agent-ready39 (Sep 14 '26)49 (Sep 15 '26)

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

Agent surface uptime llms.txt 100% (30d, checked every 6h since Sep 8 '26)