Record and replay AI agent execution for debugging
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
We haven't yet run this listing's install command through our automated sandbox check. This isn't a red flag β we're steadily working through the catalog.
π‘ Paste the JSON block into your client's configuration file under mcpServers, then restart the application.
MCP server for agent session recording and replay β debug non-deterministic agent behavior with session comparison and divergence detection.
Record every action an agent takes, replay sessions step by step, diff two runs to find behavioral regressions, and pinpoint exactly where an agent diverged from expected output.
Add to claude_desktop_config.json:
Start recording all actions for an agent session.
| Param | Type | Default | Description |
|---|---|---|---|
agent_id | string | required | Unique agent identifier |
metadata | object | {} | Optional metadata (task, model, environment) |
Returns a session_id for use with other tools.
Stop recording and return a session summary.
| Param | Type | Description |
|---|---|---|
session_id | string | Session ID from record_session |
Returns: action count, total duration, action type breakdown.
Log a single action during a recording session.
| Param | Type | Default | Description |
|---|---|---|---|
session_id | string | required | Active session ID |
action_type | string | required | Type (tool_call, llm_response, decision, error) |
input | any | required | Input to the action |
output | any | required | Output from the action |
reasoning | string | "" | Agent reasoning for this step |
duration_ms | number | 0 | Action duration in milliseconds |
Replay a recorded session step by step with full action detail.
| Param | Type | Description |
|---|---|---|
session_id | string | Session ID to replay |
Returns: complete action sequence with timing, reasoning, inputs, and outputs.
Behavioral diff between two sessions. Aligns actions by step index and highlights differences.
| Param | Type | Description |
|---|---|---|
session_id_1 | string | First session |
session_id_2 | string | Second session |
Returns: similarity ratio, identical/divergent step counts, first divergence step, and per-step diffs.
Find where an agent first deviated from expected output.
| Param | Type | Description |
|---|---|---|
session_id | string | Session to analyze |
expected_output | any | Expected final output, or array of per-step expected outputs |
If expected_output is an array, compares step by step. If a single value, finds the last matching output and flags the next step as the divergence point.
Export a session for sharing and offline analysis.
| Param | Type | Default | Description |
|---|---|---|---|
session_id | string | required | Session to export |
format | string | "json" | "json" or "markdown" |
Markdown format produces a readable transcript with step headers, reasoning, and code blocks.
| URI | Description |
|---|---|
agent-replay://sessions | All recorded sessions with status and action counts |
MIT
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/agent-replay)<a href="https://allmcps.com/mcp/agent-replay"><img src="https://allmcps.com/api/badge/agent-replay?style=directory" alt="Agent Replay on AllMCPs" /></a>