Check infrastructure health, manage incidents, and run runbooks in Faultline.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
π‘ Paste the JSON block into your client's configuration file under mcpServers, then restart the application.
Inspect callable tools, capabilities, and parameters exposed to AI agents by Faultline.
list_servicesMonitor inventory with current status (optional status filter)
get_serviceOne service + its 10 most recent checks (for diagnosis)
list_incidentsOpen incidents (or `status: "resolved"` for history)
get_incidentFull incident record: timeline, AI summary, post-mortem
acknowledge_incidentAcknowledge an incident β stops further escalation
resolve_incidentResolve with an optional note (recorded on the timeline)
Faultline is infrastructure monitoring and incident
management for DevOps/SRE teams. This repo documents faultline-mcp β
Faultline's remote MCP server, which lets AI agents (Claude, Claude Code,
Claude Desktop, or anything MCP-compatible) operate Faultline: check
infrastructure health, inspect and act on incidents, look up who's on call,
and run approved runbooks.
This repo is documentation only. The server is hosted by Faultline at
https://mcp.fltln.io/mcp (Streamable HTTP transport) β there's nothing to
install or run yourself.
flt_....https://mcp.fltln.io/mcp, sending the key as
either X-API-Key: flt_... or Authorization: Bearer flt_....Claude Code:
Clients that take raw JSON config (Claude Desktop, etc.):
| Tool | What it does |
|---|---|
list_services | Monitor inventory with current status (optional status filter) |
get_service | One service + its 10 most recent checks (for diagnosis) |
list_incidents | Open incidents (or status: "resolved" for history) |
get_incident | Full incident record: timeline, AI summary, post-mortem |
acknowledge_incident | Acknowledge an incident β stops further escalation |
resolve_incident | Resolve with an optional note (recorded on the timeline) |
who_is_on_call | Current on-call per schedule, with shift end time |
list_anomalies | Recent learned-baseline latency anomalies (observed vs baseline, z-score, hours sustained, auto-opened incident if any) |
diagnose_incident | Recommend the next action (run runbook / escalate / resolve / wait) + candidate runbooks. Analysis only β changes nothing |
run_runbook | Execute one chosen runbook against an incident β mutates infrastructure (can restart/scale services) |
run_runbook is the only tool that changes
infrastructure. Its description instructs the calling agent to use it only
after diagnose_incident recommended it and you've explicitly confirmed.Questions or issues: support@fltln.io or the Faultline dashboard.
Factual signals from GitHub, npm, and our automated checks β not a rating.
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/faultline)<a href="https://allmcps.com/mcp/faultline"><img src="https://allmcps.com/api/badge/faultline?style=directory" alt="Faultline on AllMCPs" /></a>