Skip to content

Rank #3 of 6 in AI Research Agents

FutureHouse Platform logo

FutureHouse Platform

Built-in AI assistant

FutureHouse · commercial

no public signals

Verified integrations

No integration evidence found in our corpus for this product yet — that means none was found, never that it doesn’t integrate.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

31.1/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

0.0/100

Collaboration sharing — stories about collaboration sharing in this arenaCollaboration sharingevidence →

Stories about collaboration sharing in this arena

0.0/100

Literature workflow — stories about literature workflow in this arenaLiterature workflowevidence →

Stories about literature workflow in this arena

12.0/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

8.6/100

Pricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limitsevidence →

Free-tier ceilings, usage caps, and rate limits before you have to pay

39.3/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

0.0/100

Report output — stories about report output in this arenaReport outputevidence →

Stories about report output in this arena

27.0/100

Research depth — stories about research depth in this arenaResearch depthevidence →

Stories about research depth in this arena

50.0/100

Source quality — stories about source quality in this arenaSource qualityevidence →

Stories about source quality in this arena

61.4/100

Story verdicts — every judged story with its evidenceStory verdicts

What’s free: 1 free · 0 paid · 0 enterprise · 16 not stated in evidence

?

Sorted by importance (agentic first) (high → low) · 43/43 stories · click a row’s chevron for the rationale and evidence

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full8/10C

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full7/10T

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3none±0/10

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3n/a±untestednone yet

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full7/10T

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial±7/10C

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full7/10T

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full±7/10C

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial±6/10T

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none±0/10

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1n/auntestednone yet

Pose a research question and get an autonomous multi-step investigation, not just a single-pass summary C

Agent runs

researcherResearch depth — stories about research depth in this arenaResearch depth3full8/10C

See citations for every substantive claim so I can verify it against the underlying source C

Citations

researcherSource quality — stories about source quality in this arenaSource quality3full7/10C

Get a structured report with sections, tables, and a summary that I can share with stakeholders C

Reports

analystReport output — stories about report output in this arenaReport output3partial6/10C

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3n/auntestednone yet

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3noneuntestednone yet

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3noneuntestednone yet

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3n/auntestednone yet

Search scholarly literature and primary sources, not just the open web C

Corpus

researcherSource quality — stories about source quality in this arenaSource quality2full8/10C

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial5/10T

Run a systematic screening and extraction workflow across many papers with consistent criteria C

Reviews

researcherLiterature workflow — stories about literature workflow in this arenaLiterature workflow2partial5/10T

See where sources agree and disagree instead of a single unqualified answer C

Synthesis

analystSource quality — stories about source quality in this arenaSource quality2partial5/10C

Start a long research job that keeps working unattended and notifies me when the result is ready C

Agent runs

analystResearch depth — stories about research depth in this arenaResearch depth2partial5/10C

Understand plan pricing and usage limits before committing G

Pricing

researcherPricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits2partial4/10C

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2none0/10

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2none0/10

Upload my own PDFs or corpus and have the agent research over them C

Corpus

researcherLiterature workflow — stories about literature workflow in this arenaLiterature workflow2none0/10

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2noneuntestednone yet

Share a research session or report with collaborators who can view or build on it C

Sharing

analystCollaboration sharing — stories about collaboration sharing in this arenaCollaboration sharing2noneuntestednone yet

Try the product meaningfully on a free tier or trial G

Pricing

researcherPricing limits — free-tier ceilings, usage caps, and rate limits before you have to payPricing limits1fullfree7/10C

Export results to common formats, including documents, spreadsheets, and reference-manager files G

Reports

researcherReport output — stories about report output in this arenaReport output1none0/10

Steer the depth, effort, and scope of a research run before or while it executes C

Agent runs

researcherResearch depth — stories about research depth in this arenaResearch depth1none0/10

Set up standing searches or alerts that surface new relevant sources as they appear C

Alerts

researcherLiterature workflow — stories about literature workflow in this arenaLiterature workflow1noneuntestednone yet

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1n/auntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 29 stories with headroom

What would move FutureHouse Platform’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Agenticness — how well agents can access and operate the productConnect an agent via an official MCP server

    nonemoves agent-readyimpact 45

    The evidence pack only shows a Python client (edison-client) for calling FutureHouse agents via API key, plus probes confirming no OpenAPI/MCP-related endpoints were found; there is no mention of an official MCP server for connecting agents.

  2. Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave

    nonemoves PA Scoreimpact 30

    No evidence of any data export capability, open-format export, or account portability/deletion feature; documentation covers agent/task usage and API access but nothing about exporting user data or leaving with it.

  3. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    No evidence in the pack addresses data usage for AI training opt-out, data privacy controls, or any training-data policy; the docs focus entirely on product features and API usage.

  4. Agenticness — how well agents can access and operate the productSet up automations that run autonomously in the background

    nonemoves Built-in AIimpact 30

    Missing: scheduling/cron mechanism, event-driven triggers, persistent background job management, and any docs describing recurring or unattended automation setup.

  5. Agenticness — how well agents can access and operate the productUse an official CLI

    nonemoves agent-readyimpact 30

    Evidence shows only a Python client library (edison-client, installed via pip, used programmatically with client.run_tasks_until_done) rather than a command-line interface; no CLI tool, command syntax, or terminal usage is documented anywhere in the pack.

  6. Agenticness — how well agents can access and operate the productIssue scoped/least-privilege API credentials for an agent

    nonemoves agent-readyimpact 30

    Evidence shows only a single, account-wide API token creation flow with no mention of scopes, permissions, or least-privilege controls for agents; no evidence of scoped or restricted credential issuance.

  7. Agenticness — how well agents can access and operate the productSubscribe to events via webhooks

    nonemoves agent-readyimpact 30

    No mention of webhooks, event subscriptions, or callback mechanisms anywhere in the docs; the client is a polling/run-tasks style API and OpenAPI probe returned 404s, giving no evidence of webhook support.

  8. Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples

    nonemoves API qualityimpact 30

    Evidence shows only a quickstart guide with basic client code snippets, not an interactive API reference with runnable examples; probes for OpenAPI/swagger specs and doc endpoints all returned 404s, indicating no interactive reference exists.

Showing the top 8 of 29 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map3 surfaces · 17 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

2 of 10 testable claims verified · 2 contradictedintegrity 0/100

14 distinct capability claims found in FutureHouse Platform’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

2

Verified

6

Unverified

2

Contradicted

9

Undersold

Verified (4)
Unverified (7)
Contradicted (2)
Undersold (9)
Claims outside our story set (2)

Real capability claims found in FutureHouse Platform’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Precedent tool searches across fields to determine if a research idea has been tried before and identify gaps

    source ↗
  • Molecules agent specializes in chemistry-focused molecular design and analysis

    source ↗
Suggest a story for these →

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score19 (Sep 14 '26)19 (Sep 16 '26)
Agent-ready31 (Sep 14 '26)31 (Sep 16 '26)

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

Agent surface uptime llms.txt 100% (30d, checked every 6h since Sep 8 '26)