monitoring-mcp-server
A read-only MCP server for AWS observability that provides tools to monitor EC2, EKS, RDS, ElastiCache Redis, and CloudWatch.
README
monitoring-mcp-server
A read-only MCP server for AWS observability — EC2, EKS, RDS, ElastiCache Redis, and CloudWatch. Built as a self-contained alternative when the official AWS MCP servers aren't permitted. No create, modify, or delete calls exist anywhere in it.
Tools (19)
| Area | Tools |
|---|---|
| EC2 | list_ec2_instances, get_ec2_instance_metrics, get_ec2_instance_status |
| EKS | list_eks_clusters, list_eks_nodegroups, get_eks_cluster_metrics, get_eks_namespace_metrics, get_eks_service_metrics |
| RDS | list_rds_instances, list_rds_clusters, get_rds_instance_metrics |
| Redis | list_redis_clusters, list_redis_replication_groups, get_redis_metrics |
| CloudWatch | get_active_alarms, describe_alarms, get_alarm_history, get_metric_data, list_metrics |
Design
One API call per health snapshot. Metrics use GetMetricData, which carries up
to 500 queries per request — so get_rds_instance_metrics pulls 17 metrics in a
single call instead of 17 round trips.
Adaptive periods. CloudWatch drops fine-grained data as it ages (60s → 15 days, 300s → 63 days, 3600s → 455 days). The server picks the smallest valid period for your window and caps the result at ~180 datapoints, so a 30-day query returns a readable series instead of an empty one.
Explicit gaps. A metric with no data comes back with a note explaining why
rather than being silently dropped — so "healthy but idle" is distinguishable from
"metric not published for this instance family."
Actionable errors. Botocore exceptions map to guidance: expired SSO tokens, missing permissions, throttling, wrong region, and unknown resource ids each get a specific message instead of a traceback.
Typed responses. All tools return Pydantic models, so tool output shape is stable.
Per-call overrides. Every tool accepts region and profile_name; metric tools
accept window_minutes, period_seconds, and include_datapoints.
Requirements
- Python 3.10+
- Credentials via the standard AWS chain (env vars, shared config, SSO, or IRSA).
- Read-only IAM:
ec2:Describe*,eks:List*/eks:Describe*,rds:Describe*,elasticache:Describe*,cloudwatch:GetMetricData,cloudwatch:ListMetrics,cloudwatch:DescribeAlarms,cloudwatch:DescribeAlarmHistory.ReadOnlyAccessorCloudWatchReadOnlyAccess+ service read policies cover it.
EKS note: Container Insights must be enabled for
ContainerInsightsmetrics. Without it, inventory tools still work and metric tools tell you it's disabled.
Install
cd monitoring-mcp-server
python3 -m venv .venv && source .venv/bin/activate
pip install -e .
export AWS_REGION=ap-south-1
export AWS_PROFILE=monitoring
Verify end to end (starts the server, connects as a real MCP client, calls a tool):
python test_client.py ap-south-1
Client config
Claude Desktop — ~/Library/Application Support/Claude/claude_desktop_config.json:
{
"mcpServers": {
"monitoring": {
"command": "/absolute/path/to/monitoring-mcp-server/.venv/bin/python",
"args": ["-m", "monitoring_mcp_server.server"],
"env": { "AWS_PROFILE": "monitoring", "AWS_REGION": "ap-south-1" }
}
}
}
VS Code — .vscode/mcp.json uses servers and needs "type": "stdio"
(see the file at the repo root).
Fully quit and reopen the client after editing. Changes to server code also require a client restart — the client launches the process at startup.
Example prompts
- "List stopped EC2 instances in ap-south-1."
- "Show CPU and status checks for i-0d243be7034f86ee0 over the last 6 hours."
- "Which alarms are firing, and has the API latency alarm been flapping today?"
- "Node and pod memory utilization on the prod EKS cluster."
- "RDS connections and freeable memory for orders-db over the last day."
- "Redis engine CPU, hit rate, and evictions for orders-cache-001."
Layout
monitoring_mcp_server/
aws_client.py # client factory, profile support, error translation
metrics.py # GetMetricData engine, adaptive periods, MetricSpec catalogs
models.py # Pydantic response models
ec2_tools.py # EC2 inventory + metrics + status checks
eks_tools.py # EKS clusters, node groups, Container Insights
rds_tools.py # RDS instances, clusters, metrics
redis_tools.py # ElastiCache Redis nodes, replication groups, metrics
alarms_tools.py # alarms, alarm history, generic metric access
server.py # FastMCP entry point
test_client.py # raw stdio MCP client for verification
推荐服务器
Baidu Map
百度地图核心API现已全面兼容MCP协议,是国内首家兼容MCP协议的地图服务商。
Playwright MCP Server
一个模型上下文协议服务器,它使大型语言模型能够通过结构化的可访问性快照与网页进行交互,而无需视觉模型或屏幕截图。
Magic Component Platform (MCP)
一个由人工智能驱动的工具,可以从自然语言描述生成现代化的用户界面组件,并与流行的集成开发环境(IDE)集成,从而简化用户界面开发流程。
Audiense Insights MCP Server
通过模型上下文协议启用与 Audiense Insights 账户的交互,从而促进营销洞察和受众数据的提取和分析,包括人口统计信息、行为和影响者互动。
VeyraX
一个单一的 MCP 工具,连接你所有喜爱的工具:Gmail、日历以及其他 40 多个工具。
graphlit-mcp-server
模型上下文协议 (MCP) 服务器实现了 MCP 客户端与 Graphlit 服务之间的集成。 除了网络爬取之外,还可以将任何内容(从 Slack 到 Gmail 再到播客订阅源)导入到 Graphlit 项目中,然后从 MCP 客户端检索相关内容。
Kagi MCP Server
一个 MCP 服务器,集成了 Kagi 搜索功能和 Claude AI,使 Claude 能够在回答需要最新信息的问题时执行实时网络搜索。
e2b-mcp-server
使用 MCP 通过 e2b 运行代码。
Neon MCP Server
用于与 Neon 管理 API 和数据库交互的 MCP 服务器
Exa MCP Server
模型上下文协议(MCP)服务器允许像 Claude 这样的 AI 助手使用 Exa AI 搜索 API 进行网络搜索。这种设置允许 AI 模型以安全和受控的方式获取实时的网络信息。