[Comparison · updated October 1, 2026 Read the study ](https://arxiv.org/abs/2602.08316) 

# supermemory vs Mem0

Mem0 compared itself to us with its own numbers. Here are someone else’s: an independent study that ran both under the same agent, and what teams found when they switched.

[Build with supermemory ](https://console.supermemory.ai/) [Migrate from Mem0](https://supermemory.ai/compare/supermemory-vs-mem0/#switching)

Checked October 1, 2026supermemory![](https://supermemory.ai/logos/compare/mem0.svg)Mem0

Independent study30.30% resolved24.24%, below no memory

Failing tests fixed55.95%20.24%

Right fix retrieved59.60%39.39%

Cost per task$0.58$0.62

LongMemEval-S97 · GPT-4o94.4

Open source

Managed cloud

Self-hostingVPC, data center, air-gapped

Free plan$5 of credits a month

Paid plans from$19 a month$19 a month

User profilesOne call per namespaceOne call per user

Memory versioningUpdates, extends, derivesUpdate, expiry, Dream

MCP and plugins

ComplianceSOC 2 · HIPAA · GDPRSOC 2 Type I · HIPAA · GDPR

Benchmark rows from SWE-ContextBench (arXiv 2602.08316), run by its authors. LongMemEval-S figures are each vendor’s own; [here is how to read them](https://supermemory.ai/compare/supermemory-vs-mem0/#questions).

## Benchmarks

In an independent coding study, supermemory resolved more tasks, fixed more failing tests, retrieved more past fixes, and cost less per task than Mem0\. The LongMemEval scores come from each company’s own run.

### Tasks resolved

Related bugs fixed outright by one coding agent.

SWE-ContextBench, T4Dark line: the agent without memory

supermemory

30.30%

![](https://supermemory.ai/logos/compare/mem0.svg)Mem0

24.24%

### Failing tests fixed

Tests that failed before the patch and pass after it.

SWE-ContextBench, T4Dark line: the agent without memory

supermemory

55.95%

![](https://supermemory.ai/logos/compare/mem0.svg)Mem0

20.24%

### Right past fix retrieved

How often memory brought back the fix that solved a related bug.

SWE-ContextBench, T5

supermemory

59.60%

![](https://supermemory.ai/logos/compare/mem0.svg)Mem0

39.39%

### Cost per task

What one task cost to run, memory included. Lower is better.

SWE-ContextBench, T4Dark line: the agent without memory

supermemory

$0.58

![](https://supermemory.ai/logos/compare/mem0.svg)Mem0

$0.62

### LongMemEval-S

Recall, updates and time across 500 long chat histories.

Vendor-reported

supermemory_GPT-4o_

97

![](https://supermemory.ai/logos/compare/mem0.svg)Mem0_Self-reported_

94.4

## In an independent coding study, the agent solved the most tasks with supermemory.

On 99 related tasks, the agent resolved 30 with supermemory, compared with 24 with Mem0 and 26 with no memory. These results describe the specific setups tested in the [study](https://arxiv.org/abs/2602.08316).

### Tasks resolved

[Read the paper ](https://arxiv.org/abs/2602.08316)

30.30%

resolved with supermemory 4 more tasks than no memory

1. supermemory
2. No memory
3. Mem0

### Every measure in the paper

8 of 10measures go to supermemory

supermemory ahead **8**Mem0 ahead **2**

* Accuracy
* Bugs fully fixedResolved

30.30%24.24%

+6.1 pts
* Failing tests now passingFail-to-pass, by test

55.95%20.24%

2.8× more
* Passing tests left intactPass-to-pass, by test

90.23%82.44%

+7.8 pts
* Patches that appliedPatch applies cleanly

89.90%85.86%

+4.0 pts
* Tasks with every failing test fixedFail-to-pass, by task

40.40%39.39%

+1.0 pt
* Tasks with nothing brokenPass-to-pass, by task

88.89%86.87%

+2.0 pts
* Retrieval
* Right past fix retrievedMatched anywhere in the results

59.60%39.39%

+20.2 pts
* Right fix ranked firstMatched at the top result

23.23%25.25%  
Mem0 +2.0
* Cost and time, lower is better
* Cost per taskNo memory: $0.79  
$0.58$0.62  
6% less
* Time per taskNo memory: 6.4 min

5.0 min4.7 min  
Mem0, 19 s

## Mem0 cites one of our LongMemEval runs. Here are all three.

Our report scores 97% with GPT-4o, 85.2% with Gemini 3 Pro, and 84.6% with GPT-5\. Mem0’s page compares its 94.4% run with our Gemini 3 Pro result. Pick a model to see the category results.

### LongMemEval-S, by category

[Read the report ](https://supermemory.ai/research/longmembench/)

97%

overall, 500 questions

GPT-4o answers in this supermemory run.

_0_ _25_ _50_ _75_ _100_
1. single-session-user
2. single-session-assistant
3. single-session-preference
4. knowledge-update
5. temporal-reasoning
6. multi-session

## Teams that switched felt it in production.

Benchmarks are an afternoon. These are the people who run memory every day, and what it does there.

![](https://supermemory.ai/brand/media/meadow-poster.webp)Moved from Mem0

> “Mem0 was not great. Glad to have found Supermemory.”

Scira moved its research memory off Mem0\. Indexing became reliable, recall got faster, and usage grew about 32% after launch.

![](https://supermemory.ai/brand/founders/zaid-mukaddam.jpg)Zaid MukaddamFounder, Scira[Read the story](https://supermemory.ai/blog/why-scira-ai-switched/)

40–50% fewer tokens

> “We just ditched RAG completely and went memory only through supermemory.”

![](https://supermemory.ai/brand/founders/saasjesus.jpg)Armin DaryabegiFounder, Chatarmin

No memory issues since

> “Tried almost everything. The only thing that works reliably is supermemory.”

![](https://supermemory.ai/brand/founders/harshilmathur.jpg)Harshil MathurFounder, Razorpay

Recall latencyper query, on production traffic

|         | Server  | End to end |
| ------- | ------- | ---------- |
| Search  | 187_ms_ | 356_ms_    |
| Profile | 166_ms_ | 248_ms_    |

Tokens processedper month1T+

Organizationstens of millions of end users100k+

## Signs you’ve outgrown Mem0

If any of these sound familiar, test both on your workload. Each answer links to the study or SDK example behind it.

* Memory isn’t helping your agent finish more tasks.  
30 of 99 tasks resolved, against 24 with Mem0 and 26 with no memory.[The study](https://supermemory.ai/compare/supermemory-vs-mem0/#study)
* Your agent fixes bugs it has fixed before.  
It brings back the right past fix 59.6% of the time. Mem0 does 39.4%.[The study](https://supermemory.ai/compare/supermemory-vs-mem0/#study)
* Your context spans users, projects and documents.  
Use one namespace per user or project for profiles, memories and documents.[The SDK](https://supermemory.ai/compare/supermemory-vs-mem0/#switching)
* Memory adds to what every task costs.  
$0.58 a task with supermemory, $0.62 with Mem0, in the same study.[Benchmarks](https://supermemory.ai/compare/supermemory-vs-mem0/#benchmarks)

## Keep the workflow. Change the adapter.

Use each user ID as a namespace to scope their documents, memories and profile. Both SDKs cover add, search and profile; switching means adapting payloads and responses, then backfilling existing data.

### Add. Search. Profile.

TypeScript SDKs · supermemory 5.0.1 · mem0ai 3.3.1

```
import { Supermemory } from "supermemory";const client = new Supermemory(); await client.add(userId, { content });const { results } = await client.search(userId, { query: q });const { profile } = await client.profile(userId);
```

Set `SUPERMEMORY_API_KEY` for `new Supermemory()`; pass your Mem0 key as `apiKey`. Ingestion is asynchronous: these calls show the interfaces, not immediate read-after-write. Mem0 profiles require `status: "succeeded"`; Supermemory's default dynamic memory formation can take minutes.

## Questions

Why trust a comparison written by supermemory?

You don’t have to. The main study is someone else’s, linked above, and every Mem0 number here was published by Mem0 or by that study. Where Mem0 comes out ahead, we show it.

Mem0 reports 94.4 on LongMemEval and you report 97\. Who is right?

They are different runs. Mem0’s figure comes from its own evaluation; ours from our report, with GPT-4o answering. Our Gemini 3 Pro and GPT-5 runs score 85.2 and 84.6\. Neither has been reproduced under a matched setup, which is why the independent coding study leads this page.

Do you publish LoCoMo and BEAM results?

Not in a form we can link here yet, so those rows are left out rather than estimated. When we publish them, they will be on this page with their method.

Can I run supermemory in my own infrastructure?

Yes. It runs in our cloud, in your VPC, in your own data center, and air-gapped. Talk to us about your deployment requirements.

Can I keep my user IDs when switching?

Yes. Use each user ID as a supermemory namespace, the boundary that scopes that user’s documents, memories and profile. Both hosted SDKs support user profiles, but their payloads and response shapes differ. Update your adapter and backfill existing data; the snippets show the calls, not a drop-in migration.

Sources: SWE-ContextBench, Zhu et al., arXiv 2602.08316 v3, Tables 4 and 5, run by its authors. LongMemEval-S from supermemory.ai/research; Mem0’s figure from mem0.ai/compare/mem0-vs-supermemory (July 8, 2026). Scira’s story from our case study. Production figures from our own telemetry. Mem0 is a trademark of its owner; supermemory is not affiliated with it.

## Run your own comparison.

Start free, replay your own evaluation set through both, and keep the one that remembers.

[Start building ](https://console.supermemory.ai/)
