benchmarks · reproducible · offline

Honest benchmarks

The AI-memory market publishes leaderboard numbers that don't survive independent reproduction. We do it differently: every Mnem figure on this page comes from an open harness you can run yourself — offline, in one command.

Our rules

  1. Loading…

Mnem, measured open harness

CategoryItemsRecall@1Recall@5
Loading results…

Reproduce it: python benchmarks/run_benchmarks.py — dataset, scorer, and config ship with the app's benchmark kit. No network access required.

What competitors claim vs. what reproductions found

These are not our measurements. Claims are the vendors' own self-reported figures; reproductions are third-party runs (some by competing vendors — noted where known). We list both because, as of 2026, cross-vendor leaderboard numbers are harness-relative and not comparable.

The columns that decide product fit

Where memories liveRecall pathPortability

Competitor claims and reproductions cited as of the date in each entry. Nothing on this page is hand-edited: figures are generated from harness output.

Run the numbers yourself.

Sub-10ms recall on your machine, with nothing leaving it.