Skip to content

Rank #5 of 6 in Observability & Monitoring

npm 7.3M/wkpypi 8.3M/wknpm/wk -3Mpypi/wk -159.4k

Showcase

Datadog homepage screenshot
homepage · captured Sep 2026 · view live ↗
Datadog docs screenshot
docs · captured Sep 2026 · view live ↗

Try itExperimental

See what an agent can do with Datadog before you ever sign up. Pick a story: recorded sessions replay real probe-harness transcripts; sandboxed self-drive sessions are designed and gated (docs/TRY-IT.md).

$datadog-ci versionrecorded session — replayed, not live
recorded 2026-09-05 · exit 0 · captured verbatim by our probe harness, secrets redacted

Verified integrations

Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

38.1/100

Ai assist — stories about ai assist in this arenaAi assistevidence →

Stories about ai assist in this arena

48.0/100

Alerting slos — stories about alerting slos in this arenaAlerting slosevidence →

Stories about alerting slos in this arena

60.0/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

19.5/100

Cost sampling — stories about cost sampling in this arenaCost samplingevidence →

Stories about cost sampling in this arena

11.5/100

Dashboards as code — stories about dashboards as code in this arenaDashboards as codeevidence →

Stories about dashboards as code in this arena

32.4/100

Deployment openness — stories about deployment openness in this arenaDeployment opennessevidence →

Stories about deployment openness in this arena

0.0/100

Incident response — stories about incident response in this arenaIncident responseevidence →

Stories about incident response in this arena

36.0/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

7.2/100

Otel standards — stories about otel standards in this arenaOtel standardsevidence →

Stories about otel standards in this arena

56.4/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

0.0/100

Query analytics — stories about query analytics in this arenaQuery analyticsevidence →

Stories about query analytics in this arena

24.0/100

Telemetry unified — stories about telemetry unified in this arenaTelemetry unifiedevidence →

Stories about telemetry unified in this arena

80.0/100

Story verdicts — every judged story with its evidenceStory verdicts

?

Sorted by importance (agentic first) (high → low) · 54/54 stories · click a row’s chevron for the rationale and evidence

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full9/10T

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full8/10T

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3partial5/10C

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none0/10

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full7/10C

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial6/10T

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10T

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10T

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10C

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial4/10T

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial4/10C

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1noneuntestednone yet

Alert on any telemetry signal with routing, grouping, and silencing of notifications C

Alerting

sreAlerting slos — stories about alerting slos in this arenaAlerting slos3full8/10X

Collect metrics, logs, and traces in one platform and pivot between them with shared context C

Signals

sreTelemetry unified — stories about telemetry unified in this arenaTelemetry unified3full8/10X

Have an external agent query metrics, logs, and traces through documented APIs to debug production C

Agent integration

ai-native userAi assist — stories about ai assist in this arenaAi assist3full8/10T

Send telemetry directly over OTLP with first-class OpenTelemetry support C

Otel

developerOtel standards — stories about otel standards in this arenaOtel standards3full7/10C

Analyze telemetry ad hoc with a documented query language C

Query language

developerQuery analytics — stories about query analytics in this arenaQuery analytics3partial6/10X

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3partial6/10C

Have the platform's AI investigate an alert or error and propose a probable root cause C

Ai investigation

ai-native userAi assist — stories about ai assist in this arenaAi assist3partial6/10T

Define dashboards and alerts as code (JSON models, Terraform, or API) and provision them repeatably C

As code

developerDashboards as code — stories about dashboards as code in this arenaDashboards as code3partial5/10T

See what my observability spend is, attribute it to teams or services, and catch usage spikes before the bill C

Cost

sreCost sampling — stories about cost sampling in this arenaCost sampling3disputed5/10D

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3none0/10

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3noneuntestednone yet

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3noneuntestednone yet

Instrument hosts, containers, Kubernetes, and cloud services through vendor-maintained agents and integrations C

Instrumentation

sreTelemetry unified — stories about telemetry unified in this arenaTelemetry unified2full9/10X

Define SLOs with error budgets and burn-rate alerts C

Slos

sreAlerting slos — stories about alerting slos in this arenaAlerting slos2full8/10C

Jump from a trace span to its correlated logs and metrics to debug a request end to end C

Correlation

developerTelemetry unified — stories about telemetry unified in this arenaTelemetry unified2full7/10X

Build shareable dashboards with rich visualization types and template variables C

Dashboards

sreDashboards as code — stories about dashboards as code in this arenaDashboards as code2partial6/10X

Correlate regressions with deploys and configuration changes via release or change tracking C

Change tracking

developerIncident response — stories about incident response in this arenaIncident response2partial6/10X

Declare and track incidents with timelines, on-call schedules, and escalation policies C

Incidents

sreIncident response — stories about incident response in this arenaIncident response2partial6/10C

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial6/10T

Get AI-generated summaries of incidents and alert context for responders C

Ai investigation

ai-native userAi assist — stories about ai assist in this arenaAi assist2partial6/10C

Instrument once with open standards and switch backends without re-instrumenting my code C

Otel

sreOtel standards — stories about otel standards in this arenaOtel standards2partial6/10C

Ask questions of my telemetry in natural language and get a real query or chart back C

Ai querying

ai-native userAi assist — stories about ai assist in this arenaAi assist2partial5/10T

Group and filter by high-cardinality fields (user id, request id) without pre-aggregating or defining indexes first C

Analysis

developerQuery analytics — stories about query analytics in this arenaQuery analytics2partial5/10X

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2partial4/10T

Control trace/log sampling and retention tiers to manage data volume deliberately C

Sampling

developerCost sampling — stories about cost sampling in this arenaCost sampling2partial2/10C

See application errors grouped into issues with stack traces, release tracking, and regression detection C

Errors

developerQuery analytics — stories about query analytics in this arenaQuery analytics2none0/10

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Point alert notifications at webhooks that trigger automated remediation or agents G

Alert automation

ai-native userAlerting slos — stories about alerting slos in this arenaAlerting slos2noneuntestednone yet

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2noneuntestednone yet

Run the full observability stack self-hosted in production with documented architecture and upgrade path G

Self host

sreDeployment openness — stories about deployment openness in this arenaDeployment openness2noneuntestednone yet

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2noneuntestednone yet

Enable anomaly or outlier detection that surfaces problems without hand-written thresholds C

Alerting

sreAlerting slos — stories about alerting slos in this arenaAlerting slos1full8/10C

Predict costs from transparent published per-signal pricing without talking to sales G

Cost

sreCost sampling — stories about cost sampling in this arenaCost sampling1none0/10

Spin up a local or dev instance of the platform to test instrumentation and dashboards C

Local dev

developerDeployment openness — stories about deployment openness in this arenaDeployment openness1noneuntestednone yet

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1noneuntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 40 stories with headroom

What would move Datadog’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Agenticness — how well agents can access and operate the productPlug MCP servers into this product so it can use their tools

    nonemoves agent-readyimpact 45

    Evidence only shows Datadog exposing its own official MCP server so external agents can call Datadog's tools (datadog-docs-3, datadog-probe-3) — the reverse direction of this story, which asks whether a user can plug external MCP servers into Datadog so Datadog's own AI features (e.g., Bits AI) can consume their tools.

  2. Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave

    nonemoves PA Scoreimpact 30

    Missing: documented full-account data export/backup feature, open-format export guarantees, and any evidence of successful data portability/migration by users.

  3. Openness — open source, data portability, and self-hosting storiesSelf-host the core product

    nonemoves PA Scoreimpact 30

    Datadog is a SaaS-only observability platform; no evidence of an on-premise/self-hosted core product offering exists in the pack, and its architecture (cloud dashboards, Watchdog, integrations) presumes a hosted service.

  4. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    No evidence in the pack addresses AI-training data-usage opt-out policies or controls for Datadog's own AI features (e.g., Bits AI); this is an applicable privacy-posture question for an AI-enabled product but is unaddressed by any docs or community citations.

  5. Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agent

    nonemoves agent-readyimpact 30

    Missing: documentation on restricted/scoped API keys, role-based key permissions for AI agents, or any agent-specific credential-issuance workflow.

  6. Agenticness — how well agents can access and operate the productSubscribe to events via webhooks

    nonemoves agent-readyimpact 30

    Missing: any documentation of webhook configuration, webhook payload format, or webhook-based event subscription mechanism.

  7. Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples

    nonemoves API qualityimpact 30

    Datadog has an API Reference doc page, but there's no evidence of an interactive, runnable-example reference (e.g., embedded code sandbox, try-it-now console); the OpenAPI probe even returned 404s across candidate paths, suggesting no discoverable machine-readable spec for interactive tooling.

  8. Agenticness — how well agents can access and operate the productDelegate tasks to a built-in AI assistant inside the product

    partialq5/10moves Built-in AIimpact 22.5

    Missing: detailed documentation of task-delegation capabilities and scope, and community or hands-on validation of Bits AI actually performing delegated tasks.

Showing the top 8 of 40 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map19 surfaces · 35 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

Probe proofs — replayable recordings from the probe harnessProbe proofs

Replayable recordings from our probe harness — see the Prove-It protocol to submit one.

$datadog-ci versionreproduced
$ datadog-ci version
v5.23.0
$curl -si -X POST https://mcp.datadoghq.com/api/unstable/mcp-server/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'reproduced
$ curl -si -X POST https://mcp.datadoghq.com/api/unstable/mcp-server/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'
HTTP/2 401

x-content-type-options: nosniff

strict-transport-security: max-age=31536000; includeSubDomains; preload

content-type: application/json

content-length: 27

date: Sat, 05 Sep 2026 01:11:29 GMT

{"errors":["Unauthorized"]}

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

7 of 16 testable claims verified · 2 contradictedintegrity 19/100

25 distinct capability claims found in Datadog’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

7

Verified

7

Unverified

2

Contradicted

20

Undersold

Verified (9)
Unverified (7)
Contradicted (2)
Undersold (20)
Claims outside our story set (11)

Real capability claims found in Datadog’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Static Code Analysis (SAST) scans source code for security vulnerabilities

    source ↗
  • Software Composition Analysis identifies vulnerabilities in open-source dependencies

    source ↗
  • Dynamic Instrumentation lets you add logs/traces to running code without redeploying

    source ↗
  • Continuous Profiler provides always-on code-level performance profiling

    source ↗
  • Kubernetes Autoscaling automatically scales Kubernetes workloads based on metrics

    source ↗
  • Cloud Security Posture Management scans cloud configurations for misconfigurations and risks

    source ↗
  • Sensitive Data Scanner detects and redacts sensitive data in telemetry

    source ↗
  • Audit Trail records user and system actions for compliance and auditing

    source ↗
  • Data Streams Monitoring tracks health and latency of message queues/streaming pipelines

    source ↗
  • CI Visibility tracks CI pipeline performance, test results, and flaky tests

    source ↗
  • Feature Flags feature lets teams manage and roll out feature flags

    source ↗
Suggest a story for these →

Business model

free-tiersubscription-per-seatusage-basedenterprise-custom

Modular SKU pricing: per-host infrastructure/APM subscriptions plus usage-based ingestion for logs and other signals; free tier covers up to 5 hosts of core infrastructure monitoring.

pricing ↗

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score26 (Sep 5 '26)27 (Sep 15 '26)
Agent-ready40 (Sep 5 '26)46 (Sep 15 '26)

Try Experimental

Run it in the microterminal →

Recorded agent sessions — and a live MCP handshake where the vendor ships one.

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

Agent surface uptime MCP 100% · llms.txt 100% (30d, checked every 6h since Sep 8 '26)