Record, replay, and diff AI agent tool-call workflows deterministically to detect drift and produce verifiable receipts.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent โ or use 1-click editor setup below.
We haven't yet run this listing's install command through our automated sandbox check. This isn't a red flag โ we're steadily working through the catalog.
๐ก Paste the JSON block into your client's configuration file under mcpServers, then restart the application.
Inspect callable tools, capabilities, and parameters exposed to AI agents by Reelier.
Your agents worked all night. Here's exactly what changed.
Reelier records the run that worked, freezes it as a replayable skill, and replays it deterministically โ every run comes back as a receipt: proof of what the agent did and what changed because of it. Agents make claims. Reelier writes receipts.
Agent-authored PRs (Dependabot, Claude, Codex, Cursor, โฆ) get a receipt comment in seconds: author, files changed, declared scope vs. what actually changed, sensitive paths flagged. No workflow file, no CLI, no config.
โ Install the Reelier receipts GitHub App โ free on public repos, forever.
Reelier receipt โ agent PR Author:
dependabot[bot]ยท Files changed: 2 (+119 โ41) Declared scope: none (add.reelier/scope.ymlto enable unexpected-write detection) Sensitive paths touched: โ 1 โpackage-lock.jsonProves scope and change, not correctness
A real receipt from Reelier's own repos โ see one live. Declare scope per agent in .reelier/scope.yml (or a reelier-scope block in the PR body) and the receipt reports unexpected writes. The receipt proves scope and change, never correctness or safety.
AI agents are non-deterministic โ the same prompt, a different result every run โ and they'll claim they did the work whether they did or not. Reelier records the run that worked, replays it deterministically, and writes a signed receipt that proves it. Point it at your existing CI in one workflow โ it adds a verifiable receipt, it doesn't replace your stack.
Measured on a real head-to-head benchmark, same task, same data (full method):
Deterministic replay is also ~50ร cheaper and ~59ร faster than re-running the agent, on the same benchmark.
reelier init [--dry-run] performs one checkpointed local inspection across all three Reelier paths: Path A observation coverage, Path B replay/freeze candidates, and Path C boundable/outcome-capable/shadow-only/unsupported connections and candidates. It does not deploy, gate, dispatch, upload, copy credentials, or rewrite host configuration. --dry-run writes nothing; the normal command writes only sanitized artifacts below .reelier/init/.
Teach your coding agent when to reach for Reelier. Same two commands, either host:
This installs two Agent Skills and nothing else. reelier-replay teaches your agent to freeze a
repeatable tool-call job and replay it at 0 tokens. reelier-write-safety covers bounding an
agent's writes before you grant them: what the recorder sees, what a policy refuses, and what a
receipt does and does not prove. It ships no MCP servers, so it does not wrap, observe, or gate
any tool call on its own; the reelier CLI does that, and the skills drive it via npx. Packaged in both the Agent Plugins v1.0.0 format (plugin/agent-plugins/) and the Claude Code format (plugin/claude/), generated from one source by scripts/build-plugin-packages.mjs.
Verified end to end on codex-cli 0.147.0-alpha.1.2: both formats install, enable, and the skill reaches the model. Other hosts are untested, and per-host status is tracked in docs/specs/agent-plugins-coverage-v1.md ยง4 rather than claimed here.
reelier init reveals observed coverage and local candidates without changing routes. reelier mcp --wrap "<mcp server>" proxies live tools; reelier scan/from-session freezes supported history.reelier compile turns a trace into a SKILL.md โ 0 LLM calls, minimal assertions, honest gaps printed as Open questions.reelier run replays it at Level 0 โ no LLM, byte-identical, read-only by default (writes need --allow-writes).reelier diff reports SAME or DRIFTED per step, with the failing assertion as the why โ exit 1 on drift.reelier login connects this machine to Reelier Cloud with a device code in your browser โ or set REELIER_CLOUD_URL/REELIER_CLOUD_KEY for CI and self-hosting.reelier push optionally syncs it to a ledger for a permalink and an embeddable verified-replay badge.Already have an Agent Skill? Convert it โ your skill, minus the model:
| Test | Command | Answers |
|---|---|---|
| Determinism | reelier run <skill.md> | Does this still do what it did? |
| Recovery | reelier run <skill.md> --fail N | If this broke, would the skill notice and heal? |
| Drift | reelier run <skill.md> --wrap "<your mcp server>" | Has the world moved out from under this skill? |
Taxonomy due to Mads Hansen's review of the launch post. Full semantics for each test, including recovery injection and manifest guardrails: docs/REFERENCE.md.
Dependabot and Renovate open the PR and run your test suite โ but neither knows what your agent actually does at runtime, so a dependency bump that silently changes a tool call's shape (a renamed field, a new default, a different error) sails through with green unit tests. This is the check they don't run.
Copy .github/workflows/reelier-bump-check.yml into your repo, point skill: at your own recorded .skill.md file(s), and it will: gate to PRs from dependabot[bot]/renovate[bot] (or a dependencies label), install the bumped dependency, replay your recorded skill live against it at --max-level 0 (0 tokens), and fail the check on the exact step that drifted.
This tests dependency and MCP-tool-call behavior โ it does not test model upgrades; --max-level 0 never calls an LLM. Full listing copy and setup: docs/marketplace-listing.md.
A pushed receipt carries a ladder of independently-verifiable claims โ not one blanket "verified." Depending on what you turn on, it can be signed, timestamped, CI-attested, and carry cross-checkable provider request-ids. reelier verify recomputes every claim offline, and a claim you haven't enabled just renders as an honest gap, never a shamed one.
See a real one: reelier.com/r/HWBdmGob9KeHRqXi-OEaRD0z.
Full 8-rung ladder, what each rung does and doesn't prove: docs/REFERENCE.md.
| Employee lifecycle | Reelier equivalent |
|---|---|
| Skillify a session | reelier from-session |
| Performance review | reelier run + reelier diff |
| Fleet maintenance | scheduled replays + drift alerts |
| The record | signed receipts |
"Verified" describes the record, never the agent โ a receipt proves what ran and what changed, not that the agent was good at its job.
An employment contract doesn't make an employee good โ it makes what they did visible and bounded. Same here: receipts prove scope and change, never correctness.
Factual signals from GitHub, npm, and our automated checks โ not a rating.
No reviews yet โ be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/seldonframe-reelier)<a href="https://allmcps.com/mcp/seldonframe-reelier"><img src="https://allmcps.com/api/badge/seldonframe-reelier?style=directory" alt="Reelier on AllMCPs" /></a>