Agentic AI memory with Ebbinghaus forgetting curve decay. +16pp better recall than Mem0 on LoCoMo.
Repo: sachitrafa/YourMemory
What's inside
Persistent, self-improving memory for AI agents β built on the science of how humans remember.
βΆ Try the live interactive demo Β· Website Β· Benchmarks
Every morning your AI agent treats you like a stranger. Same context re-explained. Same preferences forgotten. Every session starts from zero.
Most "memory" tools bolt a vector database onto an agent and call it done β but that's just storage. It hoards every near-duplicate until retrieval drowns in noise. A goldfish with a bigger bowl.
YourMemory is different: memory that works like a brain, not a database.
flowchart LR
A["π§ You tell your<br/>AI something"] --> B["Extract durable<br/>facts"]
B --> C["Dedup + embed<br/>+ graph-link"]
C --> D[("Memory<br/>store")]
D -->|"related facts pile up"| E["β¨ Consolidate<br/>N β 1 summary"]
D -->|"stale + unused"| F["π Decay<br/>+ prune"]
D -->|"new session"| G["β»οΈ Recall<br/>hybrid + graph"]
E --> D
G --> H["π€ Your agent<br/>picks up where<br/>it left off"]
style D fill:#0a2540,stroke:#19cdff,color:#fff
style E fill:#0c2b3a,stroke:#5eead4,color:#fff
style H fill:#0c2b3a,stroke:#19cdff,color:#fff
| Feature | What it does | |
|---|---|---|
| π§ | Consolidation | When enough related facts accumulate, they're compressed into one clean summary and the originals are archived. Memory gets sharper over time, not bloated. |
| π | Biological decay | Every memory ages on an Ebbinghaus forgetting curve. Stale, unused facts fade; important and frequently-recalled ones persist. |
| π | Entity graph | Memories link by shared people, places, and concepts β so recall surfaces what you forgot to ask for. |
| β»οΈ | Survives context resets | When the context window compacts, YourMemory hands the working context back β no re-reading files to figure out where you were. |
| π | Tamper-evident audit trail | Every read / write / delete is logged in a hash-chained ledger. Alter one record and the chain breaks. |
| π₯ | Team memory pools | Role-based shared memory, so a whole team's agents draw on the same institutional knowledge β with private memories kept private. |
| π‘οΈ | Data rights built in | One-command export (right to access) and right-to-forget (purge), plus SOC 2-aligned controls. |
| π | MCP-native & local-first | Works with Claude, Cursor, Cline, Windsurf, or any MCP client. Runs entirely on your machine β no API key, nothing leaves your system. |
One command to install. DuckDB by default (zero setup), Postgres + pgvector for teams.
Three external datasets. Every number independently reproducible β benchmark code lives in the repo. Full methodology in BENCHMARKS.md.
xychart-beta
title "Recall@5 Β· LoCoMo-10 (higher is better)"
x-axis ["Mem0", "Zep Cloud", "Supermemory", "YourMemory"]
y-axis "Recall@5 percent" 0 --> 70
bar [18, 28, 31, 59]
2Γ better recall than Zep Cloud across all 10 samples. *Supermemory and Mem0 exhausted free-tier quotas mid-benchmark; scores computed over the full 1,534 pairs.
The hardest standard benchmark for long-term memory. Each question is buried in ~53 sessions.
| Metric | Score |
|---|---|
| Recall@5 (any gold session in top-5) | 89.4% |
| Recall-all@5 (all gold sessions in top-5) | 84.8% |
| nDCG@5 (ranking quality) | 87.4% |
| System | BOTH_FOUND@5 |
|---|---|
| YourMemory (vector + BM25 + entity graph) | 71.5% |
| YourMemory (no entity edges) | 59.5% |
Entity graph edges add +12 pp β they traverse from Fact 1 to Fact 2 even when Fact 2 has low embedding similarity to the query.
Writeup: I built memory decay for AI agents using the Ebbinghaus forgetting curve
Python 3.11β3.14. No Docker, no database setup. All memory stored locally in ~/.yourmemory/.
pip install yourmemory
yourmemory-register <your-token>
yourmemory-setup
Get your token: visit yourmemoryai.xyz β enter your email β verify with a 6-digit code β copy your token.
yourmemory-setup auto-detects and wires up Claude Code, Claude Desktop, Cursor, Windsurf, and Cline, then asks which backend to use:
DATABASE_URL (needs the pgvector extension)Optional β smarter local extraction: YourMemory works out of the box with built-in heuristics. For higher-quality, fully-local fact extraction, install Ollama and
yourmemory-setuppulls the model (qwen2.5:7b, ~4.7 GB) automatically. Prefer the cloud? SetYOURMEMORY_EXTRACT_BACKEND=anthropic.
Prefer not to touch pip? Grab the standalone binary for your platform from the latest release:
| Platform | Asset |
|---|---|
| macOS (Apple Silicon) | yourmemory-macos-arm64.tar.gz |
| macOS (Intel) | yourmemory-macos-x86_64.tar.gz |
| Linux (x86-64) | yourmemory-linux-x86_64.tar.gz |
| Windows (x86-64) | yourmemory-windows-x86_64.exe.zip |
# macOS / Linux β download, extract, run
tar -xzf yourmemory-macos-arm64.tar.gz
./yourmemory-macos-arm64 register <your-token>
./yourmemory-macos-arm64 setup
./yourmemory-macos-arm64 # start the server
One executable handles every command: register, setup, ask "<question>", path, and (with no args) starts the server.
Fully self-contained & offline β the binary bundles Python, every dependency, and both ML models (the embedding model + spaCy). Nothing is downloaded on first run. The trade-off is size (~2 GB). Build your own with a single command β ./build-binary.sh β and multi-platform release binaries are produced automatically by the build workflow.
YourMemory treats memory as a living system β it grows, consolidates, forgets, and connects, the way a brain does.
Most memory tools just keep growing. YourMemory watches for clusters of related facts and, once enough accumulate, compresses them into a single clean summary β archiving the originals (never deleting, so nothing is lost).
flowchart LR
subgraph before [Related facts pile up]
A1["Railway uses Nixpacks"]
A2["Railway on Pro plan"]
A3["Railway env vars hold<br/>the Postgres URL"]
A4["Deploys on Railway<br/>with Postgres"]
end
before --> C{"cluster +<br/>LLM summarize"}
C --> S["β¨ Summary<br/>Deploys on Railway (Pro,<br/>Nixpacks) with Postgres<br/>via env vars"]
C -.->|"archived, recoverable"| ARC[("archive")]
style S fill:#0a2540,stroke:#5eead4,color:#fff
style C fill:#0c2b3a,stroke:#19cdff,color:#fff
Real example from one production store: 444 memories β 16 summaries β same knowledge, a fraction of the noise. Consolidation is event-driven (triggered when related memories pile up), not a blind nightly job.
Memory strength decays exponentially. Importance and recall frequency slow that decay:
effective_Ξ» = base_Ξ» Γ (1 β importance Γ 0.8)
strength = clamp(importance Γ e^(βeffective_Ξ» Γ active_days) Γ (1 + recall_count Γ 0.2), 0, 1)
active_days counts only days you were active β vacations don't cause memory loss. Memories below strength 0.05 are pruned automatically. Each category ages at its own rate:
| Category | Half-life | Best for |
|---|---|---|
strategy | ~38 days | Patterns that worked, architectural decisions |
fact | ~24 days | Preferences, identity, stable knowledge |
assumption | ~19 days | Inferred context, uncertain beliefs |
failure | ~11 days | Errors, wrong approaches, environment-specific issues |
Chain-aware pruning: a decayed memory is kept alive if any graph neighbour is still strong β load-bearing context survives even when rarely queried directly.
Recall runs in two rounds so it surfaces both what you asked for and what you forgot to ask for:
flowchart LR
Q["query"] --> R1["Vector + BM25<br/>hybrid search"]
R1 --> R2["Graph expansion<br/>(what you forgot to ask)"]
R2 --> S["rank by<br/>similarity Γ strength"]
S --> OUT["π― Ranked memories"]
style OUT fill:#0a2540,stroke:#19cdff,color:#fff
Subject-aware deduplication runs before every store β it embeds the subject of each sentence so "Sachit uses DuckDB" and "YourMemory uses DuckDB" stay separate (different entities), while "YourMemory uses DuckDB" and "YourMemory stores data in DuckDB" merge (same entity). No hardcoded word lists; generalises to any language.
Enterprises won't let an opaque black box store their data. So every operation β read, write, update, delete, consolidation β is appended to a hash-chained, tamper-evident audit log.
flowchart LR
E0["GENESIS"] --> E1
subgraph E1 [Event 1]
H1["row_hash =<br/>sha256(prev + data)"]
end
E1 --> E2
subgraph E2 [Event 2]
H2["row_hash =<br/>sha256(#1.hash + data)"]
end
E2 --> E3
subgraph E3 [Event 3]
H3["row_hash =<br/>sha256(#2.hash + data)"]
end
E3 --> V{"GET /audit/verify"}
V -->|chain intact| OK["β
verified"]
V -->|any row altered| BAD["β chain breaks<br/>at that row"]
style OK fill:#0a2540,stroke:#5eead4,color:#fff
style BAD fill:#3a0c14,stroke:#fb7185,color:#fff
Each row records the timestamp, actor user + agent, action, operation, target memory, source (http vs mcp), and the previous row's hash. Change any historical record and verify_chain() pinpoints exactly where the chain broke.
GET /audit # browse the trail (filter by user / action / operation)
GET /audit/verify # cryptographically verify the chain is untampered
POST /audit/prune # retention-based cleanup (90-day minimum, never lower)
Audit logging is fail-open β it never blocks a memory operation β and read/list events from the dashboard's own render loop are excluded, so the trail stays signal, not noise.
Give a whole team's agents one shared brain β without leaking anyone's private context. Memories are either shared (visible to the pool) or private (visible only to their owner).
flowchart TB
P(("π§ Team Pool<br/>shared memory"))
A["Alice's agent"] <-->|shared| P
B["Bob's agent"] <-->|shared| P
C["Carol's agent"] <-->|shared| P
A -. private .-> AP["π Alice-only"]
B -. private .-> BP["π Bob-only"]
style P fill:#0a2540,stroke:#19cdff,color:#fff
style AP fill:#0c1424,stroke:#5a6b80,color:#8294a8
style BP fill:#0c1424,stroke:#5a6b80,color:#8294a8
Role-based access is enforced per agent β what one engineer's agent learns, the whole team benefits from instantly; sensitive context stays scoped to its owner.
POST /pools # create a pool
POST /pools/{id}/members # add a member (with role)
POST /pools/{id}/memories # contribute a shared memory
POST /pools/{id}/retrieve # recall across the pool
Because memory that stores real data needs the controls to be trusted with it:
| Right | Endpoint | What it does |
|---|---|---|
| Access (DSAR export) | GET /users/{id}/export | Full export of everything stored for a user |
| Erasure (right to forget) | DELETE /users/{id}/memories | One-command purge of a user's memories |
| Portability | POST /users/{id}/import | Re-import a previous export |
| Recoverability | GET /users/{id}/archive | Retrieve consolidated-away originals |
Combined with the hash-chained audit trail and 90-day retention floor, these map directly onto the controls documented in SECURITY.md (SOC 2-aligned).
Two built-in browser UIs β no extra setup, they start automatically with the server.
http://localhost:3033/uiA full read/write view with Memories Β· Audit Β· Pools tabs: stats bar (Strong / Fading / Near-prune), per-agent tabs, memory cards with live strength bars, category filters, the audit trail, and pool management.
http://localhost:3033/graphAn interactive force-directed map of how memories connect β root memory as a bright node, neighbours color-coded by category, edge thickness = connection strength. Drag, zoom, and click any node for full content.
http://localhost:3033/graph?memoryId=42&userId=alex&depth=2
Three tools, called by your AI automatically.
| Tool | When your AI calls it | What it does |
|---|---|---|
recall_memory(query, current_path?) | Start of every task | Surfaces memories ranked by similarity Γ decay strength; spatial boost for path-matched memories |
store_memory(content, importance, category?, context_paths?) | After learning something new | Embeds, deduplicates, stores with decay; tags optional file/dir paths |
update_memory(id, new_content, importance) | When a stored fact is outdated | Re-embeds and replaces; logs the change to the audit trail |
# Store with spatial context
store_memory(
"Alex prefers tabs over spaces in Python",
importance=0.9, category="fact",
context_paths=["/projects/backend"],
)
# Next session β spatial boost fires when working in that directory
recall_memory("Python formatting", current_path="/projects/backend")
# β {"content": "Alex prefers tabs over spaces in Python", "strength": 0.87}
The only memory system that can answer questions without making any LLM API call:
yourmemory ask "what database does this project use"
# β YourMemory uses DuckDB locally and Postgres in production.
yourmemory ask "how do I fix a kubernetes deployment"
# β Not enough memory context to answer without an LLM.
When memory is strong enough it answers instantly β zero tokens, zero cloud cost, zero latency. When it isn't, it declines cleanly rather than hallucinating. Your query never leaves your machine.
MCP tools are called at the AI's discretion. The API proxy removes that uncertainty β it intercepts every LLM call, injects relevant memories automatically, and handles store_memory / update_memory with no model configuration.
Start the server (yourmemory), then point your client at localhost:3033:
from anthropic import Anthropic
client = Anthropic(
api_key="sk-ant-...",
base_url="http://localhost:3033/proxy/anthropic",
default_headers={"X-YourMemory-User": "alex"}, # per-user memory
)
# Memory is injected automatically β no other changes needed
response = client.messages.create(
model="claude-opus-4-8",
max_tokens=1024,
messages=[{"role": "user", "content": "What database do I use?"}],
)
OpenAI works identically via base_url="http://localhost:3033/proxy/openai".
flowchart LR
C["Your AI client<br/>Claude Β· Cursor Β· any MCP"] <--> Y["π§ YourMemory"]
Y --> M[("Memory<br/>store")]
Y --> A[("Audit<br/>ledger")]
style Y fill:#0a2540,stroke:#19cdff,color:#fff
style M fill:#0c1a2c,stroke:#5eead4,color:#fff
style A fill:#0c1a2c,stroke:#5eead4,color:#fff
| Component | Role |
|---|---|
| DuckDB | Default vector store β zero setup, native cosine similarity |
| PostgreSQL + pgvector | Optional β for teams or large datasets |
| NetworkX | Default graph backend (~/.yourmemory/graph.pkl) |
| Neo4j | Optional graph backend |
| sentence-transformers | Local embeddings (multi-qa-mpnet-base-dot-v1, 768 dims) |
| spaCy | Local NLP for deduplication and entity extraction |
| APScheduler | Automatic decay + pruning |
Writes hang / time out (DuckDB single-writer lock). If both the MCP server and the HTTP server run at once, they compete for the DuckDB write lock. Fix:
pkill -f yourmemory 2>/dev/null || true
rm -f ~/.yourmemory/memories.duckdb.wal ~/.yourmemory/memories.duckdb.lock 2>/dev/null || true
# restart your client
Running Claude Desktop (MCP) and Claude Code (hooks) simultaneously? Use SQLite instead β it handles concurrent readers/writers cleanly:
DATABASE_URL=sqlite:///~/.yourmemory/memories.db
PRs welcome β see CONTRIBUTORS.md.
Copyright 2026 Sachit Misra β Licensed under CC-BY-NC-4.0.
Free for personal use, education, academic research, and open-source projects. Commercial use requires a separate written agreement β mishrasachit1@gmail.com
.env.example
.github/
dependabot.yml
workflows/
build-binary.yml
docker-publish.yml
security-scan.yml
.gitignore
agents/
coding_agent.md
research_agent.md
review_agent.md
apple-touch-icon.png
assets/
logo.png
avatar.svg
benchmarks/
BENCHMARKS.md
__init__.py
beam_baseline.py
docker-compose.zep.yml
fever_contradiction.py
hotpotqa_reasoning.py
locomo_4way_results.json
locomo_4way.py
locomo_fullstack.py
locomo_mpnet.py
locomo_qa_model.py
locomo_real.py
locomo_supermemory.py
locomo_temporal.py
locomo_zep.py
locomo.py
longmemeval_fullstack.py
longmemeval_official.py
longmemeval_temporal.py
precisionmembench.py
results/
ablation_decay_20260504_091931.json
ablation_decay_20260504_091931.md
fever_20260506_091203.json
fever_20260506_155813.json
fever_20260506_161032.json
hotpotqa_20260506_091500.json
hotpotqa_20260506_092821.json
hotpotqa_20260506_094208.json
hotpotqa_20260506_121234.json
hotpotqa_20260506_160234.json
locomo_temporal.json
longmemeval_temporal.json
precisionmembench_retrieval.json
precisionmembench_session.json
precisionmembench_summary.json
run_all.py
stale_memory.py
token_efficiency.py
two_session_comparison.py
workflow_comparison.py
binary_entry.py
build-binary.sh
CLAUDE.md
cloudflare-worker/
.gitignore
package-lock.json
package.json
src/
index.ts
tsconfig.json
wrangler.toml
CONTRIBUTORS.md
demo_gifs/
temporal_boost.gif
demo.gif
docker-compose.yml
Dockerfile
docs/
policies/
01-information-security-policy.md
02-access-control-policy.md
03-incident-response-plan.md
04-vulnerability-management-policy.md
05-vendor-management-policy.md
06-data-retention-deletion-policy.md
07-change-management-policy.md
08-business-continuity-dr-policy.md
09-risk-assessment-policy.md
README.md
favicon-512.png
favicon.png
favicon.svg
glama.json
index.html
LICENSE
llms.txt
logo.svg
logo.svg.png
main.py
make_temporal_gif.py
marketing/
hero/
hero_landscape_1200x627.png
hero_square_1080x1080.png
support.js
YourMemory-Hero.dc.html
interactive.html
ph-gallery/
ph_01-hook_1270x760.png
ph_02-consolidation_1270x760.png
ph_03-context-recovery_1270x760.png
ph_04-audit-trail_1270x760.png
ph_05-shared-pools_1270x760.png
ph_06-get-started_1270x760.png
YourMemory-PH-Gallery.dc.html
memory_mcp.py
MEMORY_RULES.md
og-image.png
og-image.svg
pyproject.toml
railway.toml
README.md
requirements.txt
sample_CLAUDE.md
scripts/
reembed.py
setup_db.sh
SECURITY.md
server.json
setup.sh
sitemap.xml
SOC2_READINESS_REPORT.md
src/
__init__.py
app.py
db/
connection.py
duckdb_schema.sql
migrate.py
schema.sql
sqlite_schema.sql
graph/
__init__.py
backend.py
graph_store.py
neo4j_backend.py
networkx_backend.py
svo_extract.py
hook_templates/
__init__.py
yourmemory_observe.py
yourmemory_recall.py
yourmemory_recall.sh
yourmemory_server.py
yourmemory_session_start.py
yourmemory_store.py
yourmemory_user.sh
jobs/
decay_job.py
routes/
__init__.py
agents.py
audit.py
compact.py
dsar.py
graph_viz.py
memories.py
pools.py
proxy.py
retrieve.py
ui.py
services/
__init__.py
agent_registry.py
api_keys.py
audit.py
auth.py
compaction.py
decay.py
embed.py
extract_fallback.py
extract.py
hook_sync.py
resolve_fallback.py
resolve.py
retrieve.py
session.py
temporal.py
utils.py
support.js
tests/
test_auth.py
test_features.py
test_hook_sync.py
test_pools.py
test_store_hook.py
worker/
valtown_endpoint.ts
yourmemory_run.py
yourmemory.specFAQ
yourmemory is a Claude Code plugin with hand-picked skills for data work, indexed on Flowy. Install it with the command on its page. Its skills do not fire on their own yet. Request auto-invocation to have Flowy route them as you prompt. Free and open source.