Skip to content

Rank #4 of 6 in AI Research Agents

Undermind logo

Undermind

YC S24

Undermind · commercial

no public signals

Try itExperimental

See what an agent can do with Undermind before you ever sign up. Pick a story: recorded sessions replay real probe-harness transcripts; the live MCP handshake runs real requests from our edge, right now — including, where the server allows it, one real read-only tool call (bring your own key for auth-gated servers); sandboxed self-drive sessions are designed and gated (docs/TRY-IT.md).

$mcp-probe → https://mcp.undermind.ai/mcplive — run just now from our edge
$ press ▶ run to send one JSON-RPC initialize from our edge
real JSON-RPC against the vendor’s documented MCP endpoint — read-only, nothing is written

Verified integrations

Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

19.7/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

18.0/100

Collaboration sharing — stories about collaboration sharing in this arenaCollaboration sharingevidence →

Stories about collaboration sharing in this arena

0.0/100

Literature workflow — stories about literature workflow in this arenaLiterature workflowevidence →

Stories about literature workflow in this arena

21.6/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

12.0/100

Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limitsevidence →

Free-tier ceilings, usage caps, and rate limits before you have to pay

4.0/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

0.0/100

Report output — stories about report output in this arenaReport outputevidence →

Stories about report output in this arena

22.5/100

Research depth — stories about research depth in this arenaResearch depthevidence →

Stories about research depth in this arena

50.0/100

Source quality — stories about source quality in this arenaSource qualityevidence →

Stories about source quality in this arena

52.9/100

Story verdicts — every judged story with its evidenceStory verdicts

What’s free: 0 free · 0 paid · 4 enterprise · 12 not stated in evidence

?

Sorted by importance (agentic first) (high → low) · 43/43 stories · click a row’s chevron for the rationale and evidence

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full8/10T

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3partialenterprise5/10T

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none±0/10

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none0/10

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10X

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial3/10C

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partialenterprise2/10T

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none±0/10

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none±0/10

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1n/auntestednone yet

Pose a research question and get an autonomous multi-step investigation, not just a single-pass summary C

Agent runs

researcherResearch depth — stories about research depth in this arenaResearch depth3full8/10X

See citations for every substantive claim so I can verify it against the underlying source C

Citations

researcherSource quality — stories about source quality in this arenaSource quality3full7/10X

Get a structured report with sections, tables, and a summary that I can share with stakeholders C

Reports

analystReport output — stories about report output in this arenaReport output3partial5/10C

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3partial3/10C

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3noneuntestednone yet

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3noneuntestednone yet

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3n/auntestednone yet

Search scholarly literature and primary sources, not just the open web C

Corpus

researcherSource quality — stories about source quality in this arenaSource quality2full8/10X

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2partialenterprise6/10C

Run a systematic screening and extraction workflow across many papers with consistent criteria C

Reviews

researcherLiterature workflow — stories about literature workflow in this arenaLiterature workflow2partial6/10X

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partialenterprise5/10T

Start a long research job that keeps working unattended and notifies me when the result is ready C

Agent runs

analystResearch depth — stories about research depth in this arenaResearch depth2partial5/10X

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2none0/10

See where sources agree and disagree instead of a single unqualified answer C

Synthesis

analystSource quality — stories about source quality in this arenaSource quality2none0/10

Understand plan pricing and usage limits before committing G

Pricing

researcherPricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits2none0/10

Upload my own PDFs or corpus and have the agent research over them C

Corpus

researcherLiterature workflow — stories about literature workflow in this arenaLiterature workflow2none0/10

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2n/auntestednone yet

Share a research session or report with collaborators who can view or build on it C

Sharing

analystCollaboration sharing — stories about collaboration sharing in this arenaCollaboration sharing2noneuntestednone yet

Set up standing searches or alerts that surface new relevant sources as they appear C

Alerts

researcherLiterature workflow — stories about literature workflow in this arenaLiterature workflow1partial6/10C

Try the product meaningfully on a free tier or trial G

Pricing

researcherPricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits1disputed4/10D

Steer the depth, effort, and scope of a research run before or while it executes C

Agent runs

researcherResearch depth — stories about research depth in this arenaResearch depth1none0/10

Export results to common formats, including documents, spreadsheets, and reference-manager files G

Reports

researcherReport output — stories about report output in this arenaReport output1noneuntestednone yet

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1n/auntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 32 stories with headroom

What would move Undermind’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Agenticness — how well agents can access and operate the productDelegate tasks to a built-in AI assistant inside the product

    nonemoves Built-in AIimpact 45

    Missing: any evidence of a native, built-in AI assistant/chat agent inside Undermind's own UI that a user can delegate tasks to.

  2. Agenticness — how well agents can access and operate the productPlug MCP servers into this product so it can use their tools

    nonemoves agent-readyimpact 45

    All MCP-related evidence describes Undermind acting as an MCP *server* that other clients (Cursor, VS Code, Claude, ChatGPT) can plug into to use Undermind's own tools — the reverse of this story, which asks whether a user can plug external MCP servers into Undermind so it can use their tools.

  3. Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave

    nonemoves PA Scoreimpact 30

    No evidence of any data export feature or open-format export capability for user data/papers/workspaces; probes for docs/API endpoints also 404.

  4. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    No evidence in the pack addresses data privacy, opt-out of AI training, or data usage policies for Undermind; nothing here confirms or denies such a control exists.

  5. Agenticness — how well agents can access and operate the productPoint an agent at llms.txt or agent-oriented docs

    nonemoves agent-readyimpact 30

    Direct probes show no llms.txt, no docs.md, and no openapi spec (all 404), meaning there is no agent-consumable documentation file for a generic AI agent to fetch.

  6. Agenticness — how well agents can access and operate the productUse an official CLI

    nonemoves agent-readyimpact 30

    Evidence shows an MCP server, API access, and web/ChatGPT app integrations, but there is no mention of an official CLI tool for Undermind anywhere in the docs or probes; llms.txt, docs-md, and openapi probes all 404, and no CLI is documented.

  7. Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agent

    nonemoves agent-readyimpact 30

    No evidence of scoped/least-privilege API credential issuance for agents; the API is only mentioned generically ('Programmatic queries via API') with no docs on credential scoping, permissions, or key management, and OpenAPI probes returned 404s.

  8. Agenticness — how well agents can access and operate the productBuild against official SDKs

    nonemoves agent-readyimpact 30

    Evidence only mentions a vague 'Programmatic queries via API' for enterprise customers and an MCP server, but no official SDKs (client libraries, language bindings) are documented; probes for OpenAPI specs and docs (llms.txt, mcp.md, openapi.json) all return 404, indicating no public developer SDK resources exist.

Showing the top 8 of 32 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map6 surfaces · 17 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

7 of 10 testable claims verified · 1 contradictedintegrity 50/100

11 distinct capability claims found in Undermind’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

7

Verified

2

Unverified

1

Contradicted

7

Undersold

Verified (7)
Unverified (2)
Contradicted (1)
Undersold (7)
Claims outside our story set (2)

Real capability claims found in Undermind’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Lets users curate papers into a folder for long-term reference

    source ↗
  • Lets users star important papers across the workspace

    source ↗
Suggest a story for these →

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score19 (Sep 14 '26)16 (Sep 16 '26)
Agent-ready17 (Sep 14 '26)17 (Sep 16 '26)

Try Experimental

Run it in the microterminal →

Recorded agent sessions — and a live MCP handshake where the vendor ships one.

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

Agent surface uptime MCP 100% (30d, checked every 6h since Sep 8 '26)