Skip to content

Rank #4 of 6 in AI Memory Layers

Zep logo

Zep Software, Inc. · commercial

30.9k14.7k/yrnpm 81.3k/wk +262npm/wk -40.7kpypi/wk +10.7k

Access

Install

pippip install zep-cloud
npmnpm install @getzep/zep-cloud
brewbrew tap getzep/zepctl https://github.com/getzep/zepctl.git && brew install zepctl

Compare head-to-head

Alternatives to Zep

Showcase

Zep homepage screenshot
homepage · captured Sep 2026 · view live ↗
Zep docs screenshot
docs · captured Sep 2026 · view live ↗

Try itExperimental

See what an agent can do with Zep before you ever sign up. Pick a story: recorded sessions replay real probe-harness transcripts; sandboxed self-drive sessions are designed and gated (docs/TRY-IT.md).

$curl -si -X POST https://help.getzep.com/_mcp/server -H 'Content-Type: application/json' -d '<jsonrpc initialize>'recorded session — replayed, not live
recorded 2026-09-05 · exit 0 · captured verbatim by our probe harness, secrets redacted

Verified integrations

Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

38.4/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

9.6/100

Data lifecycle — stories about data lifecycle in this arenaData lifecycleevidence →

Stories about data lifecycle in this arena

46.0/100

Deployment self host — stories about deployment self host in this arenaDeployment self hostevidence →

Stories about deployment self host in this arena

25.5/100

Graph entity memory — stories about graph entity memory in this arenaGraph entity memoryevidence →

Stories about graph entity memory in this arena

85.0/100

Memory recall quality — stories about memory recall quality in this arenaMemory recall qualityevidence →

Stories about memory recall quality in this arena

56.3/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

30.6/100

Pricing plans — plan structure and value — what each tier costs and what it unlocksPricing plansevidence →

Plan structure and value — what each tier costs and what it unlocks

0.0/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

24.4/100

Retrieval performance — stories about retrieval performance in this arenaRetrieval performanceevidence →

Stories about retrieval performance in this arena

15.0/100

Sdk integrations — stories about sdk integrations in this arenaSdk integrationsevidence →

Stories about sdk integrations in this arena

41.1/100

Session context — stories about session context in this arenaSession contextevidence →

Stories about session context in this arena

73.8/100

Tenancy permissions — stories about tenancy permissions in this arenaTenancy permissionsevidence →

Stories about tenancy permissions in this arena

52.0/100

Story verdicts — every judged story with its evidenceStory verdicts

?

Sorted by importance (agentic first) (high → low) · 57/57 stories · click a row’s chevron for the rationale and evidence

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full9/10T

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full9/10T

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3n/a0/10

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3noneuntestednone yet

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full9/10T

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial6/10T

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial6/10T

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10C

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial5/10T

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1noneuntestednone yet

Add memories from conversations and retrieve them later with semantic search, so context persists across sessions C

Core memory

developerMemory recall quality — stories about memory recall quality in this arenaMemory recall quality3full9/10T

Get summaries of past sessions or threads so an agent can pick up where the last conversation left off C

Context assembly

developerSession context — stories about session context in this arenaSession context3full9/10C

Store memories as a knowledge graph of entities and relationships so multi-hop and entity-centric questions are answerable C

Knowledge graph

ml-engineerGraph entity memory — stories about graph entity memory in this arenaGraph entity memory3full9/10T

My agent can manage its own memory mid-conversation — adding, searching, updating, and deleting memories through tools or API calls it invokes itself C

Agent memory

ai-native userMemory recall quality — stories about memory recall quality in this arenaMemory recall quality3partial7/10T

Scope memories per user, agent, or application so one tenant's memories never leak into another's retrieval C

Isolation

developerTenancy permissions — stories about tenancy permissions in this arenaTenancy permissions3full7/10C

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3partial5/10T

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3partial4/10T

Self-host the memory layer from open-source code (e.g. via Docker) on infrastructure I control C

Self host

platform-engineerDeployment self host — stories about deployment self host in this arenaDeployment self host3partial4/10T

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3noneuntestednone yet

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3noneuntestednone yet

Connect off-the-shelf assistants (Claude, ChatGPT, Cursor) to the same memory so every tool I use shares what it knows about me C

Agent memory

ai-native userSdk integrations — stories about sdk integrations in this arenaSdk integrations2full9/10T

Delete a user's memories on demand — single memory, per-entity, or full erasure — to satisfy privacy requirements C

Forgetting

platform-engineerData lifecycle — stories about data lifecycle in this arenaData lifecycle2full9/10C

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2full8/10C

Ingest documents, JSON, and business data into memory — not just chat transcripts C

Ingestion

developerSession context — stories about session context in this arenaSession context2full8/10C

Rely on the memory layer to update, supersede, or merge memories when new information contradicts what was stored C

Core memory

developerMemory recall quality — stories about memory recall quality in this arenaMemory recall quality2full8/10X

Retrieve a token-budgeted, prompt-ready context block assembled from relevant memories in one call C

Context assembly

developerSession context — stories about session context in this arenaSession context2full8/10X

The memory layer decides for itself what is worth remembering — extracting salient facts from raw conversation and consolidating them in the background C

Agent memory

ai-native userMemory recall quality — stories about memory recall quality in this arenaMemory recall quality2full8/10X

Track when facts became valid or invalid (temporal reasoning) so the memory distinguishes current from outdated information C

Knowledge graph

ml-engineerGraph entity memory — stories about graph entity memory in this arenaGraph entity memory2full8/10T

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial7/10C

Drop the memory layer into agent frameworks like LangChain, LangGraph, CrewAI, or the Vercel AI SDK via documented first-party integrations C

Frameworks

developerSdk integrations — stories about sdk integrations in this arenaSdk integrations2partial6/10C

Govern who and what can read or write memory with roles, policies, or access-control lists, and audit that access C

Governance

platform-engineerTenancy permissions — stories about tenancy permissions in this arenaTenancy permissions2partial6/10C

Steer retrieval with metadata filters, keyword/hybrid search modes, or reranking instead of accepting a single fixed similarity search C

Retrieval controls

developerMemory recall quality — stories about memory recall quality in this arenaMemory recall quality2partial6/10C

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2partial5/10C

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial5/10T

See documented retrieval-latency targets or measured numbers (e.g. p50/p95) backing the product's speed claims C

Latency

platform-engineerRetrieval performance — stories about retrieval performance in this arenaRetrieval performance2partial5/10C

Export memories in a machine-readable format so the memory store is portable and not a lock-in trap C

Portability

platform-engineerData lifecycle — stories about data lifecycle in this arenaData lifecycle2partial4/10T

Make memories expire or decay — via TTL, expiration dates, or recency weighting — so stale facts stop surfacing C

Forgetting

developerData lifecycle — stories about data lifecycle in this arenaData lifecycle2partial4/10C

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2partial4/10C

Build against official SDKs in at least Python and TypeScript with equivalent memory APIs C

Sdks

developerSdk integrations — stories about sdk integrations in this arenaSdk integrations2partial3/10C

See published memory-quality benchmark results (e.g. LongMemEval, LoCoMo) backing the product's recall-accuracy claims C

Benchmarks

ml-engineerMemory recall quality — stories about memory recall quality in this arenaMemory recall quality2none0/10

See published pricing with a free tier and per-unit rates so I can project memory costs before committing G

Pricing

platform-engineerPricing plans — plan structure and value — what each tier costs and what it unlocksPricing plans2none0/10

Ingest at scale with async or batch processing and check the status of background memory operations C

Scale

platform-engineerRetrieval performance — stories about retrieval performance in this arenaRetrieval performance2noneuntestednone yet

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2n/auntestednone yet

Customize the memory schema — entity types, edge types, or ontology — to match my domain C

Schema customization

ml-engineerGraph entity memory — stories about graph entity memory in this arenaGraph entity memory1full8/10C

Run the memory layer fully locally — embedded in-process or against local models — without any cloud dependency C

Self host

developerDeployment self host — stories about deployment self host in this arenaDeployment self host1partial5/10T

Share selected memory across multiple agents or users (team or group memory) while keeping private memory private C

Sharing

developerTenancy permissions — stories about tenancy permissions in this arenaTenancy permissions1partial5/10C

Store images, PDFs, or other files as memory inputs and recall information from them later C

Ingestion

developerSession context — stories about session context in this arenaSession context1none0/10

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1n/auntestednone yet

Wire memory into real-time voice pipelines (e.g. LiveKit, Pipecat, ElevenLabs) with documented integrations fast enough for live conversation C

Frameworks

developerSdk integrations — stories about sdk integrations in this arenaSdk integrations1noneuntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 36 stories with headroom

What would move Zep’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Agenticness — how well agents can access and operate the productDelegate tasks to a built-in AI assistant inside the product

    nonemoves Built-in AIimpact 45

    The axis applies to this product kind (peer products hold positive or none verdicts on this story), so lack of evidence for an applicable capability is "none", never "na".

  2. Automation depth — how much of the product can run unattendedDefine rules that trigger actions automatically on events

    nonemoves PA Scoreimpact 30

    Zep is a memory/context-graph layer with search, retrieval, MCP access, and governance policies, but no evidence describes a rules engine or event-trigger mechanism that automatically fires actions on defined conditions/events.

  3. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    The evidence pack covers RTBF/deletion, access control, and MCP/CLI tooling, but contains no statement about Zep's or its LLM providers' use of customer data for model training, nor any opt-out/no-training guarantee.

  4. Agenticness — how well agents can access and operate the productSet up automations that run autonomously in the background

    nonemoves Built-in AIimpact 30

    The axis applies to this product kind (peer products hold positive or none verdicts on this story), so lack of evidence for an applicable capability is "none", never "na".

  5. Agenticness — how well agents can access and operate the productSubscribe to events via webhooks

    nonemoves agent-readyimpact 30

    No evidence in the pack mentions webhooks or event subscription mechanisms; Zep's docs cover MCP servers, CLI, SDKs, and API access but nothing about outbound event notifications or webhook subscriptions.

  6. Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples

    nonemoves API qualityimpact 30

    Zep's docs pages show static code snippets (e.g., zep-docs-21, zep-docs-29) but there is no evidence of an interactive, runnable API reference (e.g., embedded sandbox, 'try it' console, Postman/Swagger integration) anywhere in the evidence pack.

  7. Agenticness — how well agents can access and operate the productDownload a machine-readable API spec (OpenAPI or equivalent)

    nonemoves API qualityimpact 30

    Zep is an API-first service with SDKs, a CLI (zepctl), and MCP servers, so a downloadable OpenAPI spec would be a natural artifact — but no evidence pack item mentions an OpenAPI/Swagger spec, API reference export, or machine-readable schema file being available for download.

  8. Agenticness — how well agents can access and operate the productRely on versioned APIs with a documented deprecation policy

    nonemoves API qualityimpact 30

    The evidence pack covers Zep's features (memory, graph, MCP, CLI) but contains no mention of API versioning scheme or a documented deprecation policy for breaking changes.

Showing the top 8 of 36 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map19 surfaces · 39 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

Probe proofs — replayable recordings from the probe harnessProbe proofs

Replayable recordings from our probe harness — see the Prove-It protocol to submit one.

$curl -si -X POST https://help.getzep.com/_mcp/server -H 'Content-Type: application/json' -d '<jsonrpc initialize>'reproduced
$ curl -si -X POST https://help.getzep.com/_mcp/server -H 'Content-Type: application/json' -d '<jsonrpc initialize>'
HTTP/2 200

date: Sat, 05 Sep 2026 00:46:50 GMT

content-type: text/event-stream

cache-control: no-cache

content-security-policy: default-src 'self'; script-src 'self' 'unsafe-inline' 'unsafe-eval' https://app.buildwithfern.com https: blob:; style-src 'self' 'unsafe-inline' https://app.buildwithfern.com https:; img-src 'self' https://app.buildwithfern.com https: data: blob:; font-src 'self' https://app.buildwithfern.com https: data:; connect-src 'self' https://app.buildwithfern.com https: wss: ws: data: blob:; media-src 'self' https://app.buildwithfern.com https: data: blob:; object-src 'self' https://app.buildwithfern.com https: data: blob:; frame-src 'self' https://app.buildwithfern.com https: data: blob:; base-uri 'self'; form-action 'self' https://app.buildwithfern.com https:

permissions-policy: camera=(self), microphone=(self), geolocation=()

referrer-policy: strict-origin-when-cross-origin

server: cloudflare

strict-transport-security: max-age=2592000

x-content-type-options: nosniff

x-matched-path: /[domain]/api/fern-docs/mcp

x-vercel-cache: MISS

x-vercel-id: sfo1::iad1::fn5ww-1788569209940-58118193d5ad

cf-cache-status: DYNAMIC

cf-ray: a3613819fd87d438-SJC

event: message
data: {"result":{"protocolVersion":"2025-06-18","capabilities":{"tools":{"listChanged":true}},"serverInfo":{"name":"fern-docs-mcp-server","version":"1.0.0"}},"jsonrpc":"2.0","id":1}
$curl -si -X POST https://api.getzep.com/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'reproduced
$ curl -si -X POST https://api.getzep.com/mcp -H 'Content-Type: application/json' -d '<jsonrpc initialize>'
HTTP/2 401

date: Sat, 05 Sep 2026 00:46:50 GMT

content-type: text/plain; charset=utf-8

content-length: 13

www-authenticate: Bearer resource_metadata="https://api.getzep.com/.well-known/oauth-protected-resource/mcp"

x-content-type-options: nosniff

strict-transport-security: max-age=2592000

cf-cache-status: DYNAMIC

server: cloudflare

cf-ray: a361381bfc3d4c71-SJC

unauthorized
$uv run --with graphiti-core python3 -c 'from graphiti_core import Graphiti; print("PA_PROBE_OK graphiti-core imported")'reproduced
$ uv run --with graphiti-core python3 -c 'from graphiti_core import Graphiti; print("PA_PROBE_OK graphiti-core imported")'
⠋ Resolving dependencies...                                                     
⠙ Resolving dependencies...                                                     
⠋ Resolving dependencies...                                                     
⠙ Resolving dependencies...                                                     
⠙ graphiti-core==0.30.1                                                         
⠙ httpx==0.28.1                                                                 
⠙ neo4j==6.3.0                                                                  
⠙ numpy==2.5.2                                                                  
⠙ openai==3.8.0                                                                 
⠙ posthog==7.47.0                                                               
⠙ pydantic==2.13.5                                                              
⠙ pydantic-core==2.46.5                                                         
⠙ python-dotenv==1.2.3                                                          
⠙ tenacity==9.1.4                                                               
⠙ anyio==4.15.0                                                                 
⠙ certifi==2026.7.22                                                            
⠙ httpcore==1.0.9                                                               
⠙ idna==3.19                                                                    
⠙ pytz==2026.3.post1                                                            
⠙ httpx2==2.12.0                                                                
⠙ httpcore2==2.12.0                                                             
⠙ httpcore2==2.12.0                                                             
PA_PROBE_OK graphiti-core imported

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

9 of 18 testable claims verified · 0 contradictedintegrity 50/100

26 distinct capability claims found in Zep’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

9

Verified

9

Unverified

0

Contradicted

21

Undersold

Verified (13)
Unverified (13)
Undersold (21)
Claims outside our story set (1)

Real capability claims found in Zep’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Provides a guide for migrating existing memories from Mem0 to Zep

    source ↗
Suggest a story for these →

Business model

free-tierusage-basedsubscription-flatenterprise-custom

Managed cloud with a free tier, then usage-based plans priced per ingested message/data and retrieval; BYOC and Enterprise (custom OIDC, ABAC) are custom-quoted.

pricing ↗

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score32 (Sep 5 '26)29 (Sep 5 '26)
Agent-ready66 (Sep 5 '26)66 (Sep 5 '26)

Try Experimental

Run it in the microterminal →

Recorded agent sessions — and a live MCP handshake where the vendor ships one.

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

⚿ auth1 auth-gated probe

Agent surface uptime MCP 100% · llms.txt 100% (30d, checked every 6h since Sep 8 '26)