Memento

AI 与智能体

by scrypster

为 AI 工具提供持久化记忆,采用 local-first 的 knowledge graph,并支持 hybrid search。

什么是 Memento

为 AI 工具提供持久化记忆,采用 local-first 的 knowledge graph,并支持 hybrid search。

README

<p align="center"> <img src="docs/screenshots/banner.png" width="300" alt="Memento — Remember Everything, Forget Nothing" /> </p>

Memento

Give your AI tools a persistent memory — so every session starts where the last one left off.

Version Go Version License Docker MCP

Your AI starts fresh every session. Memento fixes that.

It runs on your machine, connects to any MCP-compatible AI tool, and builds a persistent knowledge graph from your conversations — entities, relationships, decisions, and context that survive every session restart.

No cloud. No API keys required. No subscriptions. Your data stays on your machine.


Quick Start

Prerequisites: Docker — or — Go 1.23+ + Node.js 18+ + Ollama

bash
git clone https://github.com/scrypster/memento.git
cd memento
./launch.sh

The script detects your environment, runs preflight checks, builds everything, and prints the exact command to connect your AI tool at the end. First run downloads Ollama models (~5 GB). After that, starts in seconds.

Your first memory

Once connected, try this in Claude:

code
"We're using PostgreSQL — chose it for pgvector support."

Close the tab. Open a new session. Ask:

code
"What database are we using?"

Your AI already knows. No re-explaining. No context window tricks.

Close the tab. Open a new session.

code
You: "What database are we using?"

→ Your AI already knows: "PostgreSQL — you chose it for pgvector support."
  No re-explaining. No context window tricks. It just remembers.

Behind the scenes, Memento built this automatically:

Graph Explorer

Every entity gets wired into a knowledge graph — people, tools, projects, decisions — with confidence scores and timestamps.


Connect Your Tools

Open http://localhost:6363/integrations — the web UI generates configs, download buttons, and connection testing for every client:

Integrations

ClientSetup
Claude CodeRun ./launch.sh — it prints the exact copy-paste command at the end. Example form: claude mcp add memento -- `pwd`/memento-mcp
Claude DesktopDownload config → drop in ~/Library/Application Support/Claude/
CursorDownload config → drop in .cursor/mcp.json + optional Cursor Rules file
WindsurfDownload config → drop in .codeium/windsurf/mcp_config.json
OpenClawAdd to ~/.openclaw/mcp.json under mcpServers — same pattern as Claude Desktop
Generic MCPAny MCP client — same pattern: command path + MEMENTO_DATA_PATH env var

The integrations page generates ready-to-paste configs with your actual binary paths and data directories. It also has connection testing, troubleshooting, and per-project workspace scoping.

Make Claude Code proactive (recommended)

The MCP connection makes tools available, but Claude won't use them automatically. Add this to ~/.claude/CLAUDE.md to make Claude store decisions and recall context without being asked:

markdown
## Memento MCP — Persistent Memory

The `memento` MCP server provides persistent cross-session memory. Use these tools proactively — don't wait to be asked.

**Store** (`store_memory`) when the user:
- States a preference or working style ("I prefer X", "always use Y format")
- Makes an architectural or technical decision
- Establishes project context that should survive session restarts
- Explicitly says "remember this" or similar

**Recall** (`recall_memory` or `find_related`) when:
- Starting a session for a known project — query for relevant context before diving in
- About to make a recommendation — check for existing preferences first
- The user asks about past decisions, choices, or "what did we decide about X"
- Something seems like it may have been discussed in a prior session

**Don't store:** transient debug output, in-progress exploration, or anything session-specific that won't matter next time.

Memories are searchable immediately after storing. Enrichment (entity/relationship extraction) runs asynchronously via local Ollama.

The web UI at Integrations → Claude Code → Make it proactive generates a version with your specific paths and connection settings, plus a download button.

See the full integration guides: Claude Code | Claude Desktop | Cursor & Windsurf | OpenClaw

Team memory — shared knowledge across your whole engineering team

Point everyone's AI tools at the same Memento instance and your team's decisions, conventions, and context become shared knowledge — queryable by anyone, attributable to anyone.

Every memory is tagged with who stored it. Memento auto-detects this from your git config, or you can set it explicitly:

bash
export MEMENTO_USER=alice   # or set in your shell profile

Or in your MCP config:

json
"env": { "MEMENTO_USER": "alice" }

Once set, you can ask:

code
What did Bob decide about the auth service this week?
recall_memory(created_by="bob", created_after="2024-01-14T00:00:00Z")

Setup: Each teammate runs Memento pointing at the same PostgreSQL database. Personal context stays personal (use a separate personal connection). Shared architectural decisions, conventions, and project context go into the shared connection.

See the team setup guide for full PostgreSQL configuration.


What Your AI Gets

Once connected, your AI has 20 tools it can call — no prompting required:

Core memory operations

ToolWhat it does
store_memoryPersist a decision or piece of context — enrichment happens async, returns in <10ms
recall_memoryRetrieve memories by ID, natural-language query, or paginated list with filters
find_relatedHybrid search: full-text + semantic vector + RRF ranking
update_memoryEdit content, tags, or metadata of an existing memory
forget_memorySoft-delete a memory (with grace period) or hard-delete permanently

Search and intelligence

ToolWhat it does
traverse_memory_graphFollow entity relationships to discover contextually connected memories (multi-hop BFS)
detect_contradictionsFind conflicting relationships, superseded-but-active memories, temporal impossibilities
explain_reasoningSurface why specific memories were retrieved for a query
get_session_context"Where did I leave off?" — recent memories grouped by topic

Memory lifecycle

ToolWhat it does
update_memory_stateMove through lifecycle: planning → active → paused / blocked / completed → archived
evolve_memoryCreate a new version that supersedes the old one — preserves full history
consolidate_memoriesLLM-assisted merge of multiple related memories into one coherent record
get_evolution_chainView the full version history of a memory from original to latest

Soft delete and recovery

ToolWhat it does
restore_memoryRecover a soft-deleted memory
list_deleted_memoriesBrowse soft-deleted memories that can still be restored
retry_enrichmentRe-run entity extraction on a memory that previously failed

Project management

ToolWhat it does
create_projectCreate a project memory with optional pre-created phases
add_project_itemAdd epics, phases, tasks, steps, or milestones under a project
get_project_treeRetrieve the full nested hierarchy of a project
list_projectsList all projects, optionally filtered by lifecycle state

Store returns in <10ms. Enrichment — entity extraction, relationship mapping, embedding generation — runs asynchronously. Your AI is never blocked.


What It Looks Like

Auto-extracted entities — zero manual input

Entities

People, projects, tools, organizations, languages, APIs — extracted automatically from your AI conversations. No tagging required.

Relationship intelligence

Relationships

Your AI knows who works_on what, which tools depend_on which services, and what the current state of each decision is — with confidence scores and timestamps.

The dashboard

Dashboard

Live enrichment queue, entity browser, relationship explorer, and graph visualizer — all in the web UI.


Why Memento

vs. Mem0

Mem0 requires cloud API keys and a paid plan for production use. Memento runs entirely on your machine with Ollama — no API keys, no cloud, no per-memory pricing. Memento also ships a full web UI with graph visualization, entity browser, and one-click integration setup. Mem0 has no web interface.

vs. Zep / Graphiti

Zep requires Neo4j or FalkorDB for its knowledge graph. Memento uses SQLite (zero deps) or PostgreSQL — no graph database to manage. Zep's open-source version is limited; the full feature set requires Zep Cloud.

vs. Built-in AI memory (ChatGPT, Claude)

Built-in memory is a flat list of facts with no relationships, no search, no graph, and no way to export or control your data. Memento gives you a structured knowledge graph you own, with hybrid search and full lifecycle management.

vs. Writing docs or wikis

Memento captures context automatically as you work — no manual effort. It builds relationships between concepts instead of isolated pages, and it's designed to be queried by LLMs, not just humans.


How It Works

code
┌─────────────────────────────────────────────────────┐
│  Your AI tool (Cursor / Claude Code / Windsurf / …) │
└─────────────────────────┬───────────────────────────┘
                          │  MCP (JSON-RPC 2.0 over stdio)
┌─────────────────────────▼───────────────────────────┐
│                   MCP Server                        │
│   store · recall · find_related · contradictions…   │
└─────────────────────────┬───────────────────────────┘
                          │
┌─────────────────────────▼───────────────────────────┐
│                Memory Engine                        │
│  ┌──────────────────────────────────────────────┐  │
│  │           Enrichment Pipeline                │  │
│  │  entity extraction → relationship mapping    │  │
│  │  → semantic embeddings → contradiction check │  │
│  └──────────────────────────────────────────────┘  │
└──────────────────┬──────────────────────────────────┘
                   │
       ┌───────────┴───────────┐
       │                       │
┌──────▼──────┐       ┌────────▼────────┐
│   SQLite    │       │  PostgreSQL     │
│  FTS5 index │       │  + pgvector     │
│  (default)  │       │  (scale-out)    │
└─────────────┘       └─────────────────┘

Features

Runs entirely offline

  • Ollama runs locally — default setup never makes an external network call
  • SQLite database is a single file you own: ~/.memento/memento.db
  • Swap to OpenAI or Anthropic when you want stronger extraction — opt-in only

Hybrid search

  • FTS5 full-text + semantic vector search fused with Reciprocal Rank Fusion (RRF)
  • Finds what you mean, not just what you typed

Knowledge graph

  • Extracts 22 entity types: people, projects, tools, languages, APIs, databases, concepts, and more
  • Maps 44 relationship types with confidence scores
  • Interactive graph explorer in the web UI

Memory lifecycle

  • Lifecycle states: planning → active → paused | blocked | completed | cancelled → archived
  • Decay scoring — stale context loses ranking weight naturally
  • Access-frequency boosting — memories you recall often stay prominent

Production-ready backends

  • SQLite (zero deps, CGo-free) for personal/local use
  • PostgreSQL + pgvector + ivfflat index for team or production deployments

Multi-connection isolation

  • Separate memory namespaces per project, client, or workspace
  • Route MCP calls to different connections with a single env var

Web UI

  • Dashboard with live enrichment queue, entity browser, relationship explorer, graph visualizer
  • One-click integration setup for every supported client
  • Connection testing, CLAUDE.md generation, Cursor Rules download
  • Tracks unrecognized LLM entity types so you can expand your taxonomy over time

LLM Providers

ProviderSetupUse when
Ollama (default)docker compose up — automaticPrivacy first, no API costs, fully offline
OpenAISet MEMENTO_LLM_PROVIDER=openai + API keyStronger extraction quality, cloud OK
AnthropicSet MEMENTO_LLM_PROVIDER=anthropic + API keyStrongest reasoning, cloud OK

Switch providers per connection — different projects can use different LLMs.


Configuration

VariableDefaultDescription
MEMENTO_PORT6363Web UI and REST API port
MEMENTO_STORAGE_ENGINEsqlitesqlite or postgres
MEMENTO_DATA_PATH./dataSQLite database directory
MEMENTO_LLM_PROVIDERollamaollama, openai, or anthropic
MEMENTO_OLLAMA_URLhttp://localhost:11434Ollama API endpoint
MEMENTO_OLLAMA_MODELqwen2.5:7bExtraction model
MEMENTO_EMBEDDING_MODELnomic-embed-textEmbedding model
MEMENTO_OPENAI_API_KEYOpenAI API key
MEMENTO_ANTHROPIC_API_KEYAnthropic API key
MEMENTO_DEFAULT_CONNECTIONDefault connection name for multi-workspace isolation
MEMENTO_CONNECTIONS_CONFIGPath to connections.json for multi-workspace setup
MEMENTO_BACKUP_ENABLEDfalseAutomated backups
MEMENTO_BACKUP_INTERVAL24hBackup frequency

PostgreSQL

bash
docker compose --profile postgres up -d
bash
MEMENTO_STORAGE_ENGINE=postgres
MEMENTO_DATABASE_URL=postgres://memento:memento_dev_password@localhost:5433/memento

Project Structure

code
memento/
├── cmd/
│   ├── memento-mcp/        # MCP server binary — connect this to your AI client
│   ├── memento-web/        # Web dashboard — entity browser, graph explorer, settings
│   └── memento-setup/      # Interactive setup wizard
├── internal/
│   ├── api/mcp/            # MCP JSON-RPC server — 20 tool handlers
│   ├── engine/             # Memory engine, enrichment pipeline, async workers
│   ├── llm/                # Ollama, OpenAI, Anthropic + circuit breaker
│   └── storage/
│       ├── sqlite/         # SQLite with FTS5 and hybrid vector search
│       └── postgres/       # PostgreSQL with pgvector and ivfflat index
├── web/
│   ├── handlers/           # HTMX handlers
│   ├── templates/          # Dashboard, graph, entities, settings, integrations
│   └── static/templates/   # MCP config snippets generated per client
├── docs/
│   └── integrations/       # Per-client integration guides
├── migrations/             # SQL schema migrations
└── docker-compose.yml

Contributing

Issues and PRs welcome. Open an issue before starting significant work.

bash
go test ./...

go build -o memento-mcp ./cmd/memento-mcp/
go build -o memento-web ./cmd/memento-web/
go build -o memento-setup ./cmd/memento-setup/

License

MIT — see LICENSE.


Built by

MJ Bonanno — software architect and founder of Scrypster.


Remember everything. Forget nothing. Unlike Leonard Shelby, your context is here to stay — searchable, versioned, and backed by a knowledge graph that never fades.

常见问题

Memento 是什么?

为 AI 工具提供持久化记忆,采用 local-first 的 knowledge graph,并支持 hybrid search。

相关 Skills

Claude接口

by anthropics

Universal
热门

面向接入 Claude API、Anthropic SDK 或 Agent SDK 的开发场景,自动识别项目语言并给出对应示例与默认配置,快速搭建 LLM 应用。

想把Claude能力接进应用或智能体,用claude-api上手快、兼容Anthropic与Agent SDK,集成路径清晰又省心

AI 与智能体
未扫描165.3k

RAG架构师

by alirezarezvani

Universal
热门

聚焦生产级RAG系统设计与优化,覆盖文档切块、检索链路、索引构建、召回评估等关键环节,适合搭建可扩展、高准确率的知识库问答与检索增强应用。

面向RAG落地,把知识库、向量检索和生成链路系统串联起来,做架构设计时更清晰,也更少踩坑。

AI 与智能体
未扫描23.5k

多智能体架构

by alirezarezvani

Universal
热门

聚焦多智能体系统架构设计,梳理 Supervisor、Swarm、分层和 Pipeline 等模式,覆盖角色定义、通信协作与性能评估,适合规划稳健可扩展的 AI agent 编排方案。

帮你系统解决多智能体应用的架构设计与协同编排难题,适合构建复杂 AI 工作流,成熟度高、社区认可也很亮眼。

AI 与智能体
未扫描23.5k

相关 MCP Server

顺序思维

编辑精选

by Anthropic

热门

Sequential Thinking 是让 AI 通过动态思维链解决复杂问题的参考服务器。

这个服务器展示了如何让 Claude 像人类一样逐步推理,适合开发者学习 MCP 的思维链实现。但注意它只是个参考示例,别指望直接用在生产环境里。

AI 与智能体
89.1k

知识图谱记忆

编辑精选

by Anthropic

热门

Memory 是一个基于本地知识图谱的持久化记忆系统,让 AI 记住长期上下文。

帮 AI 和智能体补上“记不住”的短板,用本地知识图谱沉淀长期上下文,连续对话更聪明,数据也更可控。

AI 与智能体
89.1k

by deusdata

热门

持久化的代码库知识图谱,可跨会话保留上下文,在 session 重启或上下文压缩后仍能继续使用。

专治 AI 编程助手“会话失忆”,把代码库沉淀为持久知识图谱,重启或压缩上下文后也能无缝续上开发状态。

AI 与智能体
36.7k

评论