hybrid-memory-mcp

hybrid-memory-mcp

Universal FastMCP bridge into a hybrid memory stack, enabling MCP-capable agents to save, search, read, and recall memories using OpenViking and Honcho backends.

Category
访问服务器

README

hybrid-memory-mcp

Universal FastMCP bridge into a hybrid memory stack. Any MCP-capable agent harness (Claude Code, Codex, RooCode, OpenClaw, Pi) adds one streamable-HTTP URL + bearer token and gains four tools against the SAME memory the host agent uses.

  • OpenViking (primary, read+write) — mem_save, mem_search, mem_read
  • Honcho (secondary, read-only reasoning) — mem_recall

Endpoint

URL:    http://<HOST>:8899/mcp        (bind a private LAN/VPN IP — not public)
Auth:   Authorization: Bearer <HMCP_BEARER_TOKEN from .env>
Transport: streamable-http

<HOST> is set via HMCP_BIND_HOST in .env (defaults to 127.0.0.1). Bind it to a private tailnet/LAN IP so only trusted machines on your network can reach it — this server is not designed for public internet exposure.

Tools

tool args backend
mem_save content (req), category, source_agent OV disk-write + async index
mem_search query (req), limit=8 OV /search/search
mem_read uri (req, must start viking://user/<user>/) OV /content/read
mem_recall query (req) Honcho /chat (read-only)

category enum: pattern | entity | event (default) | preference. Every saved fact is footer-tagged [via: <source_agent>] for cross-harness attribution.

Configuration (.env, chmod 0600)

HMCP_BEARER_TOKEN=<generate: python3 -c "import secrets;print('hmcp_'+secrets.token_urlsafe(32))">
HMCP_BIND_HOST=127.0.0.1        # set to your private tailnet/LAN IP
HMCP_BIND_PORT=8899
OPENVIKING_ENDPOINT=http://127.0.0.1:1933
HONCHO_BASE_URL=http://localhost:8000
HONCHO_WORKSPACE=hermes
HONCHO_PEER=<your-peer-id>

Client config

Replace <TOKEN> with the value of HMCP_BEARER_TOKEN, and <HOST> with your bind IP.

Claude Code

claude mcp add --transport http hybrid-memory http://<HOST>:8899/mcp \
  --header "Authorization: Bearer <TOKEN>"

Codex (~/.codex/config.toml)

[mcp_servers.hybrid_memory]
url = "http://<HOST>:8899/mcp"
http_headers = { Authorization = "Bearer <TOKEN>" }

Generic streamable-http (RooCode / OpenClaw / Pi / any MCP client)

{
  "mcpServers": {
    "hybrid-memory": {
      "type": "streamable-http",
      "url": "http://<HOST>:8899/mcp",
      "headers": { "Authorization": "Bearer <TOKEN>" }
    }
  }
}

Run

python3 -m venv venv && ./venv/bin/pip install -r requirements.txt
./venv/bin/python server.py     # binds $HMCP_BIND_HOST:$HMCP_BIND_PORT, mounts /mcp

For production, run under a supervisor with restart-on-failure (systemd unit, Restart=always).

Checks

bash specs/hybrid-memory-mcp/checks/roundtrip.sh    # C3 save->search->read (server stopped)
bash specs/hybrid-memory-mcp/checks/validation.sh   # C6 input-validation (server stopped)
bash specs/hybrid-memory-mcp/checks/auth.sh         # C4 bearer enforced (spins ephemeral server)
bash specs/hybrid-memory-mcp/checks/mcp_handshake.sh # C8 MCP initialize handshake
bash specs/hybrid-memory-mcp/checks/restart.sh      # C7 supervised persistence (after unit installed)

Design notes

  • Async index waiting-list. OV's content/write runs the semantic+vector index inline behind one global write lock (~40s/write). mem_save does the durable part synchronously (write the file to disk — the memory is persisted and recoverable immediately) and hands the slow index to a background worker that drains one job at a time. Saves return in <10ms; search visibility follows asynchronously.
  • Crash-durable queue. Each queued index job drops a pending-marker on disk; the worker deletes it after a successful index; on startup any survivors are re-queued. A kill -9 with items in flight loses nothing — markers reconcile on restart, and the worker also re-enqueues on transient failure so it self-heals once OV recovers.
  • OV 0.3.8 create-mode bug. content/write cannot CREATE new files, so mem_save writes the file to the local store first, then content/write mode=replace to index.
  • Graceful degradation. mem_recall returns {"status": "honcho unavailable", ...} instead of crashing when Honcho is down or slow.

推荐服务器

Baidu Map

Baidu Map

百度地图核心API现已全面兼容MCP协议,是国内首家兼容MCP协议的地图服务商。

官方
精选
JavaScript
Playwright MCP Server

Playwright MCP Server

一个模型上下文协议服务器,它使大型语言模型能够通过结构化的可访问性快照与网页进行交互,而无需视觉模型或屏幕截图。

官方
精选
TypeScript
Magic Component Platform (MCP)

Magic Component Platform (MCP)

一个由人工智能驱动的工具,可以从自然语言描述生成现代化的用户界面组件,并与流行的集成开发环境(IDE)集成,从而简化用户界面开发流程。

官方
精选
本地
TypeScript
Audiense Insights MCP Server

Audiense Insights MCP Server

通过模型上下文协议启用与 Audiense Insights 账户的交互,从而促进营销洞察和受众数据的提取和分析,包括人口统计信息、行为和影响者互动。

官方
精选
本地
TypeScript
VeyraX

VeyraX

一个单一的 MCP 工具,连接你所有喜爱的工具:Gmail、日历以及其他 40 多个工具。

官方
精选
本地
graphlit-mcp-server

graphlit-mcp-server

模型上下文协议 (MCP) 服务器实现了 MCP 客户端与 Graphlit 服务之间的集成。 除了网络爬取之外,还可以将任何内容(从 Slack 到 Gmail 再到播客订阅源)导入到 Graphlit 项目中,然后从 MCP 客户端检索相关内容。

官方
精选
TypeScript
Kagi MCP Server

Kagi MCP Server

一个 MCP 服务器,集成了 Kagi 搜索功能和 Claude AI,使 Claude 能够在回答需要最新信息的问题时执行实时网络搜索。

官方
精选
Python
e2b-mcp-server

e2b-mcp-server

使用 MCP 通过 e2b 运行代码。

官方
精选
Neon MCP Server

Neon MCP Server

用于与 Neon 管理 API 和数据库交互的 MCP 服务器

官方
精选
Exa MCP Server

Exa MCP Server

模型上下文协议(MCP)服务器允许像 Claude 这样的 AI 助手使用 Exa AI 搜索 API 进行网络搜索。这种设置允许 AI 模型以安全和受控的方式获取实时的网络信息。

官方
精选