16 Mm Memory Layer
Group 4 · Knowledge & Memory · element 16 of 58

Memory Layer

Agents that remember Tuesday.

Turns stateless calls into continuity.

Holders this quarterMem0 · Zep · Letta · Supermemory · Cognee

Vital signs

NecessityOptional
Price band$ · under $30/mo
MaturityExperimental
Editionv2026.Q3
Last verified2026-08-06

Why it's on the table

On the table, Memory Layer (Mm) is seat 16 of 58, in the Knowledge & Memory family. It is an experimental element — promising, volatile, and worth a contained experiment rather than a commitment. Budget curiosity, not dependence. It is optional: plenty of companies run without it — until a specific trigger (scale, regulation, cost, or customers) makes it essential for them. It sits in the lowest paid band — lunch money against the hours it returns.

The verdict — v2026.Q3 · verified 2026-08-06
Mem0
Mem0, for most builders — the biggest ecosystem (61.7k GitHub stars, 90k+ developers, exclusive memory provider for AWS's Agent SDK), the cheapest real entry ($19/mo), and the fastest shipping cadence (background consolidation 'Dream' landed Aug 4, 2026). Zep wins when you're an enterprise that needs temporal knowledge graphs over business data with provenance, ABAC, and SOC 2/HIPAA. Letta wins when the self-improving agent IS the product, not a feature. Supermemory wins on ingestion volume per dollar and cross-tool personal memory; Cognee when you want to own the whole pipeline open-source. The looming threat is bundling: Claude and ChatGPT now remember natively, Anthropic's developer memory tool is GA, and AWS/Google/Redis sell memory as a metered primitive — a standalone layer must earn its keep across models and apps, or it's a feature.

Memory Layer: the top 5 — v2026.Q3

Edition v2026.Q3 · ranking, pricing and status verified 2026-08-06.

  1. 1Mem0Mem0 (YC S24)

    Hobby free (10k adds/mo) · Starter $19/mo · Pro $249/mo · Enterprise custom (on-prem, SSO)

    Best for Adding 'remembers the user' to an existing app in an afternoon — a hosted extraction-and-retrieval API that works with any model and framework.

    The category's center of gravity: 61.7k GitHub stars, 90k+ registered developers, 13M+ package downloads, and API calls that grew 35M (Q1 2025) to 186M (Q3 2025). Raised $24M (seed + Series A led by Basis Set, Oct 2025) and became the exclusive memory provider for AWS's Agent SDK. Still shipping fast: Dream, its background memory-consolidation engine, launched Aug 4, 2026.

    Watch Its 'SOTA' LoCoMo claims are disputed — Zep published a rebuttal showing a corrected implementation beat Mem0's setup, and Mem0's own paper had a full-context baseline outperforming it. Retrieval quotas bite: 1k retrievals/mo free and 5k on the $19 tier push retrieval-heavy apps to $249 quickly.

    $24M raised (seed $3.9M + $20M Series A led by Basis Set); API calls 35M Q1 → 186M Q3 2025 (Oct 28, 2025) [src] · 61.7k GitHub stars, Apache-2.0; new memory algorithm shipped Apr 2026 (checked Aug 2026) [src] · Dream background memory consolidation launched Aug 4, 2026; 90k+ developers on platform [src]
  2. 2ZepZep AI

    Free 10k credits/mo · Flex $1,250/yr · Flex Plus $3,750/yr · Emerging Cos $13k/first year (SOC 2, HIPAA BAA) · Enterprise custom

    Best for Enterprises whose agents must remember evolving business facts — temporal knowledge graphs over chat plus CRM/app data, with provenance tracking and access control.

    The engineering-serious enterprise pick: temporal context graphs (open-sourced as Graphiti, 29k stars) with sub-200ms retrieval at 100M graphs, provenance lineage for synthesized facts (Jul 2026), attribute-based access control (Jul 2026), and SSO-gated memory over MCP (Jun 2026). Customers include Zscaler, Samsung, and HoneyBook; S&P Global called it a likely 'de facto partner in this layer of the enterprise agent stack.'

    Watch No cheap paid tier — the jump from free to $1,250/yr excludes hobbyists, and the self-hosted Community Edition was deprecated (code moved to legacy/), so the real product is cloud-only. Its 94.7% LoCoMo / 90.2% LongMemEval numbers are vendor-run, like everyone else's.

    94.7% LoCoMo at 155ms, 90.2% LongMemEval at 162ms; 161–168ms retrieval at 10M–100M graphs (vendor, Aug 2026) [src] · Graphiti OSS: 29k stars, v0.29.2 released Jun 8, 2026 [src] · Provenance tracking (Jul 14, 2026), ABAC (Jul 9, 2026), enterprise-SSO memory MCP server (Jun 30, 2026) [src]
  3. 3LettaLetta (ex-MemGPT, UC Berkeley)

    Free · Pro $20/mo (20 stateful agents) · Developer $0.10/active agent/mo + $0.00015/sec tool exec · Enterprise custom

    Best for Building agents whose memory and self-improvement are the product — stateful agents that learn across sessions, from the researchers who invented the pattern.

    The intellectual origin of the category: MemGPT (Oct 2023) invented LLM virtual context management, and Letta commercialized it with a $10M Felicis-led seed at $70M (Sep 2024). Its March 2026 pivot doubled down on what worked — Letta Code, a model-agnostic harness with git-backed memory files ('you own the memory, you choose the model'), plus research on sleep-time compute and continual learning that the rest of the field imitates.

    Watch Strategic churn is real: the Mar 16, 2026 'next phase' deprecated large chunks of the platform (Filesystem, server-side templates, MCP integrations, sleep-time agents, tool rules), and the original server repo is now labeled legacy. Smallest disclosed war chest of the leaders.

    $10M seed led by Felicis at $70M post (Sep 23, 2024); founders created MemGPT at Berkeley's Sky Lab [src] · Pivot to Letta Code harness with git-backed memory announced Mar 16, 2026; legacy features deprecated by mid-April [src] · 24k stars (Apache-2.0) on the letta repo, now labeled the legacy V1 API server (Aug 2026) [src]
  4. 4SupermemorySupermemory

    Free (~$5 usage) · Pro $19/mo · Max $100/mo · Scale $399/mo · Enterprise custom (self-host)

    Best for High-volume, cost-sensitive context: memory + RAG + connectors (Slack, Gmail, Drive, GitHub) in one API, plus a personal memory app that follows you across AI tools.

    The fastest riser: 1.5B+ memories stored, sub-300ms recall claims, and a shipping pace that produced a POSIX-compatible semantic filesystem (SMFS, May 28, 2026 — claimed 55% cheaper agentic retrieval), Context Cloud (May 18, 2026), and default 'dynamic dreaming' consolidation (May 25, 2026) — all on a $3M pre-seed (Susa Ventures, Oct 6, 2025). Vendor cites internal deployments at Google and Nissan.

    Watch Tiny funding versus rivals and a split focus (consumer app + infra API). Benchmark leadership claims (LongMemEval, LoCoMo, ConvoMem) and the Google/Nissan logos are vendor-sourced with no independent verification.

    1.5B+ memories saved; sub-300ms recall claimed (vendor, Aug 2026) [src] · $3M pre-seed led by Susa Ventures, with Browder Capital and SF1 (Oct 6, 2025) [src] · SMFS semantic filesystem launched May 28, 2026 claiming 55% cheaper agentic retrieval [src]
  5. 5CogneeTopoteretes

    OSS free (Apache-2.0) · Cloud: Free 1M tokens · Standard $2.50/1M tokens + $5/workspace · Enterprise custom (BYO cloud)

    Best for Teams who want to own the memory pipeline — an open-source ECL (extract-cognify-load) engine that builds combined knowledge-graph + vector memory over your own databases.

    The credible open-source alternative to hosted memory APIs: 29.8k stars, 5M+ SDK runs monthly, v1 shipped, and real production proof (Bayer runs agentic research memory on it; Knowunity POC'd 40,000 students in 2 days). Deploys self-hosted, Docker, on-prem, or cloud, and plugs into Claude Code, Cursor, LangGraph, and MCP.

    Watch The smallest commercial operation in the top 5 — funding undisclosed (Pebblebed and others, amounts unannounced) — and the ECL pipeline demands data-engineering appetite that Mem0's two-line SDK doesn't.

    29.8k GitHub stars, 5M+ SDK runs monthly, v1 released (vendor, Aug 2026) [src] · Bayer production deployment of agentic research memory; Knowunity 40k-student POC in 2 days (vendor case studies) [src] · Cloud pricing $2.50/1M tokens + $5/workspace, free 1M-token tier (Aug 2026) [src]

Memory Layer: the top 8 compared

Edition v2026.Q3 · ranking, pricing and status verified 2026-08-06.

Memory Layer — the top 8 compared. Edition v2026.Q3, verified 2026-08-06.
ToolApproachOpen sourceHosted entry priceSelf-hostLatency claimComplianceMCP / plugins
Mem0Extract + consolidate API, opt. graphYes (Apache, 61.7k★)Free · $19/moYes (OSS)70x vector-search cut (Jul 2026)SOC 2 I, HIPAA-readyOpenMemory MCP, Claude Code
ZepTemporal knowledge graphGraphiti only (29k★)Free · $1,250/yrNo (CE deprecated)sub-200ms at 100M graphsSOC 2 II, HIPAA BAASSO-gated memory MCP
LettaAgent-native context mgmt, git-backed filesYes (Apache, 24k★)Free · $20/moYesn/a (in-agent)Enterprise SSO tierLetta Code harness
SupermemoryMemory + RAG + connectors + SMFSPartialFree · $19/moEnterprise onlysub-300ms recallEnterprise self-hostClaude/Cursor plugins, MCP
CogneeECL pipeline: graph + vector + relationalYes (Apache, 29.8k★)Free · $2.50/1M tokYes (core)unpublishedEnterprise SLAsMCP, Claude Code, LangGraph
LangMemMemory SDK for LangGraphYes (MIT, 1.5k★)Free (BYO infra)Yesn/aVia LangGraph PlatformLangGraph-native
HonchoPeer modeling + insight reasoningYes (AGPL, 4.9k★)Managed api.honcho.devYes (Docker)unpublishedNone publishedAPI/SDK
AgentCore MemoryManaged events + strategies primitiveNo$0.25/1k eventsNounpublishedAWS-gradeAWS-native

How to choose your memory layer

If you have a working app and just need it to remember users across sessions, this week
Mem0 — free for 10k memory-adds a month, $19 after, two-line SDK, works with any model. Watch the retrieval quota, not the add quota.
If you're an enterprise whose agents must track facts that change over time — accounts, policies, patient state — with audit and access control
Zep. Temporal graph with provenance and ABAC is the point; budget $1,250/yr minimum, $13k if you need SOC 2 II + HIPAA BAA as a startup.
If the agent itself is the product and it must demonstrably learn and improve over weeks
Letta — the MemGPT lineage, sleep-time compute, and git-backed memory you can inspect and version. Accept the platform-pivot risk.
If you're ingesting everything — email, Slack, docs, screenshots — and cost per token retrieved decides the architecture
Supermemory: connector-heavy, sub-300ms claims, and SMFS cut agentic retrieval costs 55% by the vendor's own math. Verify their benchmarks against your data.
If one assistant, one user, one vendor — 'my chatbot should remember me'
Don't buy a layer. Claude and ChatGPT memory do this natively, and Anthropic's client-side memory tool is GA for developers. A standalone memory layer only earns its keep across models, apps, or agents.

Memory Layer: the whole field

19 more tools tracked in this category, including 6 dead, renamed, or sunsetting — a reference that hides the graveyard isn't one. Verified 2026-08-06.

Memory Layer — every tool we track, including 6 dead, renamed, or sunsetting. Edition v2026.Q3, verified 2026-08-06.
ToolMakerWhat it isEntryStatus
GraphitiZep (Apache-2.0)Temporal knowledge-graph framework under Zep's cloud — 29k stars, v0.29.2 Jun 2026; the OSS on-ramp to Zepfree + your infraactive
OpenMemoryMem0Local-first MCP memory for coding agents (Cursor, Claude Code, VS Code) — project-scoped preference recallfreeactive
LangMemLangChain (MIT)Memory SDK for LangGraph agents — only 1.5k stars and increasingly folded into LangGraph's own persistence storyfree + your infraactive
Claude memory + memory toolAnthropicNative memory in Claude (Oct 23, 2025, 559 HN points) plus a GA client-side memory tool for developers — the bundling threat in personbundledactive
ChatGPT memoryOpenAIAutomatic, continuously-updated memory replaced the manual system; consumer-side only, no developer API exposurebundledactive
AgentCore MemoryAWSMetered memory primitive in Bedrock AgentCore — $0.25/1k events, $0.75/1k records/mo stored; commoditization from above$0.25/1k eventsactive
Memory BankGoogle (Vertex AI Agent Engine)Managed long-term user memory for Vertex agents — Google's answer to the same primitiveusage-basedactive
Redis Agent Memory Server / IrisRedisOSS reference implementation (296 stars) graduated into Redis Iris, a managed agent-memory service on Redis Cloudfree OSS · cloud usageactive
MemUNevaMind AI14.2k-star lightweight memory-as-markdown-wiki across agents and devices; auto-extracts reusable skills from session logsfree + tokensactive
HonchoPlastic LabsPeer-modeling memory (AGPL, 4.9k stars) — background reasoning builds psychological representations of users; managed at api.honcho.devfree + managed tieractive
Memobasememodb.ioProfile-based long-term memory (2.7k stars) — structured user profiles + time-aware events, sub-100ms retrieval focusfree + tokensactive
MIRIXMirix AISix-type multi-agent memory (3.5k stars) with screen-observation capture — personal-assistant angle, local-firstfreeactive
HyperspellHyperspell'Company brain' context graph over 50+ SaaS sources surfaced as an agent-readable filesystem — overlaps element Kwunverifiedactive
Memarycommunity (MIT)Knowledge-graph agent memory, 2.6k stars — last release Oct 2024, momentum gonefreefading
PaprPapr AIFormer memory-API startup; site now sells AI GTM workflow apps — memory positioning quietly abandonedfading
MemGPTUC Berkeley → LettaThe Oct 2023 paper/project that started the category — renamed Letta with the Sep 2024 commercializationrenamed
Zep Community EditionZepThe self-hosted OSS memory server that built Zep's following — deprecated, code moved to legacy/; cloud-only nowdead
MotorheadMetal (getmetal)Early Rust memory/retrieval server for LLMs — unsupported since Dec 2023; the category's first gravedead
Rayriftsolo developerDeveloper-focused memory layer listed for takeover/acquisition on HN, Jan 31, 2026 — a marker of how crowded the low end gotdead

Memory Layer: the category in numbers

Edition v2026.Q3 · ranking, pricing and status verified 2026-08-06.

  • Mem0 raised $24M (Oct 28, 2025) on API-call growth from 35M (Q1 2025) to 186M (Q3 2025) and became AWS Agent SDK's exclusive memory provider [src]
  • Bundling from above: AWS AgentCore Memory meters memory at $0.25/1k events, Google ships Vertex Memory Bank, Redis launched Iris, and Anthropic's developer memory tool went GA — the primitive is commoditizing [src]
  • Native assistant memory became table stakes: Claude memory launched Oct 23, 2025 (559 HN points); ChatGPT moved to fully automatic, continuously-updated memory [src]
  • 2026's feature battleground is consolidation: Supermemory made 'dynamic dreaming' default (May 25), Mem0 shipped Dream (Aug 4), Letta published sleep-time compute research — everyone now sleeps [src]
  • Benchmark credibility crisis: Zep's rebuttal (May 2025, updated Jun 2026) showed Mem0's LoCoMo comparison used a flawed Zep implementation, and Mem0's own paper had a full-context baseline (~73%) beating its system (~68%) [src]
  • Architecture still unsettled: Letta's Mar 16, 2026 pivot moved memory from server-side databases to git-backed files and deprecated much of its platform API — the field's founder rethinking the field's premise [src]

Memory Layer: method & sources

Ranking charter: ecosystem adoption, verified traction, shipping velocity, and pricing accessibility — explicitly NOT vendor benchmark scores, because every vendor here claims to lead LoCoMo/LongMemEval and the Zep-Mem0 dispute plus LoCoMo's full-context-baseline problem make those numbers unusable for ranking. Conflicts resolved: Supermemory's raise is reported as $2.6M in some coverage and $3M on the vendor blog — we cite the vendor's own Oct 6, 2025 post ($3M, Susa-led); single-source, flagged. Mem0's site says 62,590 stars while the GitHub page showed 61.7k the same day — timing/rounding, we cite GitHub. Zep and Cognee funding amounts are undisclosed; treat their runway as unverified. Supermemory's Google/Nissan deployments and benchmark leads are vendor-claimed only. Letta's 24k-star repo is now labeled the legacy V1 server — star count overstates current-product momentum. Adjacent elements: vector databases (pgvector, Pinecone, Turbopuffer) → Vd; company wikis and 'company brain' knowledge bases (incl. Hyperspell's overlap) → Kw; meeting memory (Granola, Otter) → Mt; MCP servers as distribution → Mc. Native memory in Claude/ChatGPT is covered here only as the bundling threat, not as picks — it doesn't cross apps or models, which is this element's whole job. Ranking criteria: verified commercial traction, independent satisfaction surveys, agent benchmarks, and founder-fit (price floor, lock-in, surfaces). Editorial, never paid — the charter. Machine-readable twin: mm.json.

All sources (26)
  1. https://mem0.ai/pricing
  2. https://mem0.ai
  3. https://mem0.ai/blog
  4. https://github.com/mem0ai/mem0
  5. https://techcrunch.com/2025/10/28/mem0-raises-24m-from-yc-peak-xv-and-basis-set-to-build-the-memory-layer-for-ai-apps/
  6. https://www.getzep.com
  7. https://www.getzep.com/pricing
  8. https://blog.getzep.com
  9. https://blog.getzep.com/lies-damn-lies-statistics-is-mem0-really-sota-in-agent-memory/
  10. https://github.com/getzep/graphiti
  11. https://www.letta.com
  12. https://docs.letta.com/pricing
  13. https://www.letta.com/blog/our-next-phase
  14. https://techcrunch.com/2024/09/23/letta-one-of-uc-berkeleys-most-anticipated-ai-startups-has-just-come-out-of-stealth/
  15. https://github.com/letta-ai/letta
  16. https://supermemory.ai
  17. https://supermemory.ai/blog
  18. https://www.cognee.ai
  19. https://docs.cognee.ai
  20. https://aws.amazon.com/bedrock/agentcore/pricing/
  21. https://platform.claude.com/docs/en/agents-and-tools/tool-use/memory-tool
  22. https://www.anthropic.com/news/memory
  23. https://help.openai.com/en/articles/8590148-memory-faq
  24. https://github.com/langchain-ai/langmem
  25. https://github.com/plastic-labs/honcho
  26. https://github.com/redis/agent-memory-server

Our take

Nascent and worth watching. The first element where 'my agent knows my business' becomes literal.

Combines with

Appears in compounds

The Agent Stack

This is element 16 of 58. The table is versioned quarterly — when a tool loses its seat, the changelog records the succession.

Explore the full table →