OKF Knowledge Agent MCP Server

OKF Knowledge Agent MCP Server

An LLM-managed knowledge base following the Open Knowledge Format (OKF) v0.1 spec. Provides MCP tools: kb_query, kb_add, kb_update, kb_status over stdio or streamable HTTP.

Category
访问服务器

README

understory 🌱

Memory that grows.

The layer beneath your agents: a self-wiring, plain-markdown memory. Every fact your agents learn is filed as a markdown concept, cross-linked into a living knowledge graph, and kept healthy by the agent itself — searchable, diffable, and entirely yours. Runs great on local models.

Bundles follow the Open Knowledge Format (OKF) v0.1 spec — plain markdown files with YAML frontmatter, readable by humans, diffable in git, portable across tools.

Three ways in, one agent:

  • MCP servermemory_query / memory_add / memory_update / memory_status / memory_maintain tools over stdio or streamable HTTP. Each call drives an internal LLM agent with the OKF spec in its system prompt.
  • Web UI — browse the bundle (tree, concept viewer, update log, conformance badge), see the memory as an Obsidian-style force-directed graph (drag/pan/zoom, colored by type, sized by connections, orphans ringed red, click to open), and chat with the same agent to test it. Tool calls render inline so you can watch it work.
  • Query-path replay — every agent run (query/mutation/chat) records its traversal (searches → reads → writes) as a compact notation, persisted under <bundle>/.traces/. The graph view lists recent runs; selecting one replays the path as numbered directed hops over the graph — visited concepts ringed, search hits dotted, everything else faded.
  • CLIpnpm agent:query "..." / pnpm agent:mutate "..." smoke entries.

Design rule: conformance is enforced in code, not prompts. The deterministic bundle layer validates frontmatter (type required), regenerates index.md files, appends log.md entries (newest-first, spec §7), and sandboxes all paths to the bundle root. The LLM decides what to change; the code guarantees the result is a conformant bundle.

Quick start (Docker)

No clone needed — the image is public. Save this as docker-compose.yml:

services:
  understory:
    image: ghcr.io/thecodacus/understory:latest
    ports:
      - "3800:3800"
    volumes:
      # Your memory lives here as plain markdown — a named volume, or point
      # a bind mount (e.g. ./my-memory:/bundle) at any OKF bundle.
      - understory-memory:/bundle
    environment:
      BUNDLE_ROOT: /bundle
      # Pick ONE provider:
      # 1) Local llama.cpp / llama-swap (model auto-discovered; start llama-server with --jinja)
      LLM_PROVIDER: llamacpp
      LLAMACPP_BASE_URL: http://your-inference-box:8080
      # 2) Anthropic
      #LLM_PROVIDER: anthropic
      #ANTHROPIC_API_KEY: sk-ant-...
      # 3) OpenRouter
      #LLM_PROVIDER: openrouter
      #OPENROUTER_API_KEY: sk-or-...
    restart: unless-stopped

volumes:
  understory-memory:
docker compose up -d

Then:

  • Web UI → http://localhost:3800 — browse the memory, watch the graph, chat with the agent
  • MCP endpointhttp://localhost:3800/mcp (streamable HTTP) — register it in any MCP client:
    claude mcp add --transport http ustory http://localhost:3800/mcp
    
  • Your agent now has memory_query / memory_add / memory_update / memory_status / memory_maintain, and gets a seed overview of the memory at every session start.

Teach it something (memory_add: "We deploy on Fridays, never Mondays"), then open the graph and watch the concept wire itself in. Deploying with Portainer? Use docker-compose.portainer.yml as a repository stack.

Stack

pnpm monorepo:

Package What
packages/core OKF bundle layer (zero LLM) + agent (Vercel AI SDK tool loop: search/read/list/write/patch/delete) + provider registry
packages/server Express: MCP streamable-HTTP at /mcp, stdio bin, REST browse API at /api/*, streaming chat at /api/chat, serves the web build
packages/web Vite + React + TS + Tailwind: bundle browser + agent chat (useChat)

Providers (env-selected, swappable per chat): Anthropic (default), OpenRouter, llamacpp (llama.cpp llama-server / llama-swap — model auto-discovered from /v1/models, loaded model preferred), local (any other OpenAI-compatible endpoint).

llama.cpp

# on the inference box — --jinja enables OpenAI-style tool calling
llama-server -m model.gguf --jinja --host 0.0.0.0 --port 8080

# here — no model id needed, it's discovered
LLM_PROVIDER=llamacpp LLAMACPP_BASE_URL=http://inference-box:8080 \
BUNDLE_ROOT=./sample-bundle node packages/server/dist/index.js

Works behind llama-swap too: discovery prefers the currently loaded model so a query doesn't trigger a multi-minute model swap. Pin a specific model with LLM_MODEL=.

From source

pnpm install
pnpm build
cp .env.example .env   # add your API key

BUNDLE_ROOT=./sample-bundle ANTHROPIC_API_KEY=sk-... node packages/server/dist/index.js
# → http://localhost:3800  (web UI + /api + /mcp)

Or build the container yourself: docker compose up --build (the repo's docker-compose.yml builds from source and mounts ./sample-bundle).

Dev mode (server on :3800, Vite HMR on :5180 with proxy):

BUNDLE_ROOT=./sample-bundle pnpm --filter @understory/server dev
pnpm --filter @understory/web dev

MCP registration (Claude Code / Desktop)

claude mcp add ustory \
  -e BUNDLE_ROOT=/path/to/your/bundle \
  -e ANTHROPIC_API_KEY=sk-... \
  -- node /path/to/understory/packages/server/dist/mcp/stdio.js

Or point an HTTP MCP client at http://host:3800/mcp.

Seed memory

A client LLM that only sees four bare tool names never gets the instinct to check memory. So at session start the server injects a compact overview of what the knowledge base contains (directories, concepts with types + descriptions, recent activity) through both channels that reach the model:

  1. the MCP initialize instructions field (clients like Claude put it in the system prompt), and
  2. the memory_query tool description — the universal fallback every tool-calling client loads.

The seed regenerates fresh for every new session. After memory_add / memory_update in a long-lived (stdio) session, the tool description refreshes via tools/list_changed, so the session sees its own writes. Out-of-band edits (hand edits, other clients) are picked up on the next session.

Graph health & maintenance

Memory is a graph, not a pile of notes, and graphs rot: concepts go orphaned (nothing links to them) and links go broken. Two mechanisms keep it healthy:

  • Write-time linking — new knowledge either enriches the concept it belongs to (an attribute of an existing entity is patched in, not filed separately) or, when it's a distinct entity, is created and back-linked from related concepts. Contradictions are superseded in place, never left standing alongside the old value.
  • memory_maintain — a deterministic lint (orphans + broken links, surfaced in memory_status under graph) drives an internal agent to wire orphans into related concepts and fix dangling links. Run it periodically to counter drift; it's a no-op when the graph is already healthy.

This design mirrors the pattern in Karpathy's LLM Wiki (index.md + log.md, create-vs-enrich, lint for orphans). Deferred from that pattern until scale warrants: an explicit page-type schema, and hybrid FTS5+embedding search (the naive scan in search.ts is fine into the low thousands of concepts).

Tests

pnpm test                                  # core: 18 tests (spec §5/§6/§7/§9, sandbox, search, concurrency)
pnpm --filter @understory/server exec tsx scripts/mcp-smoke.mts   # MCP stdio round-trip (needs SMOKE_BUNDLE + an API key)

Environment

See .env.example. BUNDLE_ROOT is required; GIT_AUTOCOMMIT=true commits every mutation.

推荐服务器

Baidu Map

Baidu Map

百度地图核心API现已全面兼容MCP协议,是国内首家兼容MCP协议的地图服务商。

官方
精选
JavaScript
Playwright MCP Server

Playwright MCP Server

一个模型上下文协议服务器,它使大型语言模型能够通过结构化的可访问性快照与网页进行交互,而无需视觉模型或屏幕截图。

官方
精选
TypeScript
Magic Component Platform (MCP)

Magic Component Platform (MCP)

一个由人工智能驱动的工具,可以从自然语言描述生成现代化的用户界面组件,并与流行的集成开发环境(IDE)集成,从而简化用户界面开发流程。

官方
精选
本地
TypeScript
Audiense Insights MCP Server

Audiense Insights MCP Server

通过模型上下文协议启用与 Audiense Insights 账户的交互,从而促进营销洞察和受众数据的提取和分析,包括人口统计信息、行为和影响者互动。

官方
精选
本地
TypeScript
VeyraX

VeyraX

一个单一的 MCP 工具,连接你所有喜爱的工具:Gmail、日历以及其他 40 多个工具。

官方
精选
本地
graphlit-mcp-server

graphlit-mcp-server

模型上下文协议 (MCP) 服务器实现了 MCP 客户端与 Graphlit 服务之间的集成。 除了网络爬取之外,还可以将任何内容(从 Slack 到 Gmail 再到播客订阅源)导入到 Graphlit 项目中,然后从 MCP 客户端检索相关内容。

官方
精选
TypeScript
Kagi MCP Server

Kagi MCP Server

一个 MCP 服务器,集成了 Kagi 搜索功能和 Claude AI,使 Claude 能够在回答需要最新信息的问题时执行实时网络搜索。

官方
精选
Python
e2b-mcp-server

e2b-mcp-server

使用 MCP 通过 e2b 运行代码。

官方
精选
Neon MCP Server

Neon MCP Server

用于与 Neon 管理 API 和数据库交互的 MCP 服务器

官方
精选
Exa MCP Server

Exa MCP Server

模型上下文协议(MCP)服务器允许像 Claude 这样的 AI 助手使用 Exa AI 搜索 API 进行网络搜索。这种设置允许 AI 模型以安全和受控的方式获取实时的网络信息。

官方
精选