WorkspaceGuard
MCP server wrapping the WorkspaceGuard CLI as a single generic run tool for workspace usage checks.
README
<!-- mcp-name: io.github.RudrenduPaul/workspaceguard -->
WorkspaceGuard
<p align="center"> <a href="#install">Install</a> • <a href="#quickstart">Quickstart</a> • <a href="#cli-reference">CLI Reference</a> • <a href="#comparison">Comparison</a> • <a href="#faq">FAQ</a> </p>
Per-workspace usage metering and fail-closed quota caps for one shared self-hosted AI assistant deployment (Odysseus or a compatible backend).

Run Odysseus (or a compatible self-hosted assistant) for your household or small team and there's no way to see who sent how many messages this month, or to stop one person's usage from burning through everyone else's API budget. WorkspaceGuard is a sidecar that adds that layer: per-workspace message counts, an optional monthly cap that fails closed, and a CLI (or JSON) report an admin or another agent can read.
npx workspaceguard-cli usage
-> alex [alex@example.com]: 812 messages this period, cap 1000 (81%)
-> jordan [jordan@example.com]: 203 messages this period, cap unlimited
Install
npm install -g workspaceguard-cli
Or run it without installing:
npx workspaceguard-cli usage
The package is workspaceguard-cli; the command it installs is workspaceguard. A genuine, independent Python port with the same CLI surface and --json shapes is published separately as workspaceguard-cli on PyPI (pip install workspaceguard-cli, see python/).
Quickstart
# Register the workspaces sharing one deployment (identity = the header value
# your reverse proxy sets after authenticating, e.g. Cloudflare Access).
workspaceguard add-workspace alex --identity alex@example.com
workspaceguard add-workspace jordan --identity jordan@example.com
# Optional: cap alex at 1000 messages/month. Omit for unlimited (the default).
workspaceguard set-cap alex 1000
# See usage for every workspace.
workspaceguard usage
Real output from a fresh install:
-> alex [alex@example.com]: 0 messages this period, cap 1000 (0%)
-> jordan [jordan@example.com]: 0 messages this period, cap unlimited

Features
- Per-workspace message counting. Every request through the sidecar's
chat()entry point increments a per-workspace, per-month counter (src/core/usage.ts), isolated so one workspace's usage never leaks into another's. - Quota enforcement that fails closed. A workspace at its cap gets a
QuotaExceededErrorbefore the backend is ever called. If the usage store is corrupted or unreadable, WorkspaceGuard blocks requests instead of silently resetting everyone's count to zero (see CHANGELOG.md). - Agent-native
--jsonon every command.workspaceguard usage --jsonreturns structured output an orchestrator can parse directly, no screen-scraping. - AES-256-GCM vault with real key rotation.
workspaceguard rotate-key <id>re-encrypts a workspace's secrets under a new key and invalidates the old ciphertext. - A self-healing circuit breaker. Backend calls open the circuit after 3 consecutive failures, then retry through a half-open probe and close again on success, instead of staying tripped forever.
- One choke point, not scattered checks.
chat()insrc/core/isolation-guard.tsis the single place every request flows through: resolve workspace, check quota, call backend, record usage. - Two independent, tested distributions. The TypeScript package (npm) and the Python port (PyPI) implement the same design against separate test suites: 41/41 TypeScript tests and 50/50 Python tests passing as of this writing.
CLI reference
Every command accepts --json for a structured, agent-native output shape instead of the human-readable text shown below.
| Command | What it does |
|---|---|
workspaceguard init |
Initializes the data directory and vault for this deployment. |
workspaceguard add-workspace <id> --identity <value> |
Registers a workspace, idempotent on repeat calls for the same id. --identity is parsed positionally and must immediately follow <id>; it is not a free-standing flag. |
workspaceguard status [--json] |
Lists configured workspaces. |
workspaceguard usage [--json] |
Per-workspace message count, cap, and percent-used for the current month. |
workspaceguard set-cap <id> <count|none> |
Sets or clears a workspace's monthly message cap. |
workspaceguard rotate-key <id> |
Rotates a workspace's vault encryption key (invalidates the old ciphertext). |
workspaceguard scan [--json] |
Isolation config scan (scaffold stub, carried over from the original build; always returns an empty finding list today). |
workspaceguard -h, --help |
Prints the command list above and exits 0. |
workspaceguard -V, --version |
Prints the installed package version and exits 0. |

Global options
| Option | What it does |
|---|---|
--data-dir <path> |
Data directory for config, vault, and usage data. Takes precedence over WORKSPACEGUARD_DATA_DIR. |
--force |
init only: regenerate the master key even if an existing key file at the resolved data dir looks corrupted or truncated. |
--json |
Structured, agent-native output instead of human-readable text. |
[!WARNING]
--forcepermanently invalidates anything already encrypted under the old master key. Only use it when the existing key file is confirmed unrecoverable.
Data directory resolution, in order: --data-dir flag, then WORKSPACEGUARD_DATA_DIR env var, then ~/.workspaceguard. This used to default to the current working directory with no override -- running init from the wrong shell could silently write a live encryption key into an unrelated directory. init on an existing, valid key is idempotent (it loads and reuses that key); init on a key file that exists but doesn't decode to a valid key refuses to overwrite it without --force.
$ workspaceguard usage --json
{"ok":true,"usage":[{"workspaceId":"alex","identity":"alex@example.com","monthlyMessageCap":1000,"percentUsed":81,"period":"2026-07","messageCount":812,"estimatedBytes":48213}]}
The --json mode is what makes this agent-native rather than just human-convenient: an orchestrator or monitoring agent can call workspaceguard usage --json and parse the result directly instead of scraping terminal output.
Library API
import { createWorkspaceGuard, MockAdapter, QuotaExceededError } from "workspaceguard-cli";
const guard = await createWorkspaceGuard({ dataDir: "./data", backend: new MockAdapter() });
await guard.addWorkspace("alex", "alex@example.com");
await guard.setCap("alex", 1000);
try {
await guard.chat("alex@example.com", "hello");
} catch (err) {
if (err instanceof QuotaExceededError) {
// alex is over their monthly cap
}
}
const report = await guard.usageReport();
The Python port exposes the same shape: from workspaceguard import create_workspace_guard, MockAdapter, QuotaExceededError.
MCP Server
WorkspaceGuard's Python distribution ships a Model Context Protocol server, so an MCP-compatible agent (Claude Desktop, Claude Code, an orchestrator) can call WorkspaceGuard directly as a tool instead of shelling out to the CLI and parsing text.
pip install "workspaceguard-cli[mcp]"
It exposes one tool, run, a generic subprocess wrapper: pass it the same argument list you'd pass on the command line, and it shells out to the installed workspaceguard binary, parses the resulting JSON, and returns it. Every failure mode (missing binary, launch error, timeout, non-zero exit, unparseable output) comes back as a plain {"error": ...} dict instead of raising, so a bad call can't crash the server.
run(args=["usage", "--json"])
# -> {"ok": true, "usage": [{"workspaceId": "alex", "identity": "alex@example.com", "monthlyMessageCap": 1000, "percentUsed": 0, "period": "2026-08", "messageCount": 0, "estimatedBytes": 0}]}
To register it with an MCP-compatible client such as Claude Desktop, add it to the client's server config:
{
"mcpServers": {
"workspaceguard": {
"command": "workspaceguard-mcp"
}
}
}
This assumes workspaceguard-mcp is already on PATH (installed via the mcp extra above). If you installed it somewhere else, replace "command" with the full path to the console script.
Comparison
WorkspaceGuard is a sidecar, not a competing product. It sits in front of an Odysseus deployment (or a compatible backend) and adds the one layer that backend doesn't provide.
| Capability | WorkspaceGuard | Odysseus (native) |
|---|---|---|
| Per-user isolation (chat history, memory, API keys) | Not reimplemented; treated as already solved | Yes, built in by default |
| Per-workspace message counting | Yes | No |
| Monthly quota caps, fail-closed | Yes | No |
CLI / --json usage report |
Yes | No |
| License | MIT | AGPL-3.0 |
What is WorkspaceGuard, and why does it exist
This project originally set out to add per-user workspace isolation (separate chat history, memory, API keys) to a self-hosted AI chat platform. A feasibility spike found that Odysseus already enforces per-user ownership on chat history, memory, and API tokens by default, so building a competing isolation layer would have duplicated work Odysseus already does correctly.
WorkspaceGuard instead keeps its tested isolation engine (namespace separation, an AES-256-GCM vault with real key rotation, fail-closed identity resolution, a self-healing circuit breaker) as the identity-resolution substrate, and builds the layer Odysseus doesn't provide: usage metering and quota enforcement per workspace.
Free tier (this repo, MIT): per-workspace message counting, monthly cap enforcement, a CLI/JSON usage report. Not in this repo: a hosted, multi-tenant billing dashboard is a separate, closed-source product, mentioned here only as a roadmap item and never merged into this MIT codebase.
Architecture
src/core/isolation-guard.ts-- the single choke point (chat()) every request flows through: resolve workspace, check quota, call backend, record usage.src/core/usage.ts-- the usage-metering engine this project adds: per-workspace, per-month counters with automatic period rollover, andQuotaExceededErrorenforcement.src/core/vault.ts,src/core/namespace.ts,src/core/circuit-breaker.ts-- the original isolation-engine code, kept as the identity and workspace-boundary substrate the metering layer reads from.src/adapters/-- theBackendAdapterinterface.MockAdapteris the only implementation today; a real Odysseus HTTP adapter has not been built yet.
Backend-specific behavior never enters src/core/ directly. Everything goes through BackendAdapter.
Trust boundary
WorkspaceGuard trusts an upstream identity header (default: Cf-Access-Authenticated-User-Email) to resolve the workspace.
[!WARNING] This service must never be directly reachable from the network. Only run it behind a trusted proxy that sets that header (Cloudflare Access, Tailscale, etc.). This boundary is documented, not code-enforced.
What's real vs. not yet built
- Real and tested: usage metering, quota enforcement, the original isolation engine (vault, namespace separation, circuit breaker), and the CLI with
--jsonmode, verified by 41/41 passing TypeScript tests and 50/50 passing Python tests. - Not yet built: a real Odysseus HTTP adapter (only
MockAdapterexists today) and a hosted, multi-tenant billing dashboard (deliberately out of scope for this MIT repo).
Docs
FAQ
Q: What does WorkspaceGuard actually do?
A: It adds per-workspace usage metering and quota enforcement in front of one shared self-hosted AI assistant deployment. It counts messages per workspace per month, lets you set an optional cap that fails closed once hit, and gives you (or an agent) a workspaceguard usage report. It does not add chat history, memory, or API key isolation itself; that already exists by default in the target platform (see "What is WorkspaceGuard" above), and WorkspaceGuard's own isolation code (src/core/vault.ts, src/core/namespace.ts) is kept only as the identity-resolution substrate the metering layer reads from.
Q: What's WorkspaceGuard's actual differentiator?
A: Narrow scope done well: not a full billing platform, and not a reimplementation of isolation the backend already has. Every request flows through one choke point (chat() in src/core/isolation-guard.ts), quota enforcement fails closed on a corrupted usage store instead of silently resetting everyone's usage to zero (see CHANGELOG.md), and every command supports --json for agent-native output.
Q: How does WorkspaceGuard compare to Odysseus? A: It isn't a competing product. WorkspaceGuard is a sidecar that sits in front of an Odysseus deployment (or a compatible backend); it doesn't replace anything Odysseus already does. See the comparison table above for the specific capability split.
Q: What platforms does WorkspaceGuard run on?
A: The npm package (workspaceguard-cli) requires Node.js 20 or newer (engines.node in package.json). The Python port in python/ requires Python 3.9 through 3.13 (see the classifiers in python/pyproject.toml). Neither distribution ships a platform-specific binary, so both run wherever their respective runtime does (Linux, macOS, Windows).
Q: Is WorkspaceGuard a CLI, a library, or both?
A: Both, in both distributions. The CLI (workspaceguard <command>) covers init, add-workspace, status, usage, set-cap, rotate-key, and scan. The same functionality is importable directly (createWorkspaceGuard from the TypeScript package, create_workspace_guard from the Python package) for anything that wants to call it from code instead of shelling out.
Q: What's a real current limitation I should know about before relying on this?
A: The only backend adapter implemented today is MockAdapter, an in-memory adapter used for tests and local experimentation. A real Odysseus HTTP adapter has not been built yet (see docs/integrations/backends.md), so WorkspaceGuard does not yet forward live chat traffic to an actual Odysseus deployment. The metering and quota logic itself is real and tested; the network bridge to a live backend is the piece still outstanding.
Q: Does WorkspaceGuard need its own API keys, or hold any of my AI provider credentials?
A: No. The only backend adapter that exists right now (MockAdapter) is in-memory and calls no external API. All backend-specific behavior is isolated behind the BackendAdapter interface (src/adapters/), so WorkspaceGuard's own code never needs to see provider credentials directly.
Q: Is WorkspaceGuard free to use commercially? A: Yes. This repository is MIT licensed in full, with no dual licensing and no feature gate. The hosted, multi-tenant billing dashboard mentioned above is a separate, closed-source product described only as a roadmap item; no billing-dashboard code lives in, or is withheld from, this MIT codebase.
Contributing and security
See CONTRIBUTING.md and SECURITY.md. Notable changes are tracked in CHANGELOG.md.
License
MIT.
推荐服务器
Baidu Map
百度地图核心API现已全面兼容MCP协议,是国内首家兼容MCP协议的地图服务商。
Playwright MCP Server
一个模型上下文协议服务器,它使大型语言模型能够通过结构化的可访问性快照与网页进行交互,而无需视觉模型或屏幕截图。
Magic Component Platform (MCP)
一个由人工智能驱动的工具,可以从自然语言描述生成现代化的用户界面组件,并与流行的集成开发环境(IDE)集成,从而简化用户界面开发流程。
Audiense Insights MCP Server
通过模型上下文协议启用与 Audiense Insights 账户的交互,从而促进营销洞察和受众数据的提取和分析,包括人口统计信息、行为和影响者互动。
VeyraX
一个单一的 MCP 工具,连接你所有喜爱的工具:Gmail、日历以及其他 40 多个工具。
graphlit-mcp-server
模型上下文协议 (MCP) 服务器实现了 MCP 客户端与 Graphlit 服务之间的集成。 除了网络爬取之外,还可以将任何内容(从 Slack 到 Gmail 再到播客订阅源)导入到 Graphlit 项目中,然后从 MCP 客户端检索相关内容。
Kagi MCP Server
一个 MCP 服务器,集成了 Kagi 搜索功能和 Claude AI,使 Claude 能够在回答需要最新信息的问题时执行实时网络搜索。
e2b-mcp-server
使用 MCP 通过 e2b 运行代码。
Neon MCP Server
用于与 Neon 管理 API 和数据库交互的 MCP 服务器
Exa MCP Server
模型上下文协议(MCP)服务器允许像 Claude 这样的 AI 助手使用 Exa AI 搜索 API 进行网络搜索。这种设置允许 AI 模型以安全和受控的方式获取实时的网络信息。