The full upstream README, mirrored here for reference. Install config, tool schemas, adoption signals, and an original overview live on the Mastyf.ai listing page.
Perimeter security for your AI.
Runtime enforcement, policy control, and full audit trail for every AI action.
Website · Quick start · Policy · Dashboard · GitHub
AI agents can read your files, push code, query databases, execute shell commands, and call external APIs. They do it autonomously, at machine speed.
Traditional security controls weren't built for that.
Mastyf.ai acts as a perimeter security layer for AI. It intercepts every tool call, evaluates it against your security policies using multi-agent swarm analysis, and blocks malicious or unauthorized actions before they execute.
Every decision is enforced, logged, and auditable.
| Threat | What it looks like |
|---|---|
| Prompt injection | Malicious instructions embedded in tool arguments to hijack agent behavior |
| Path traversal | Attempts to access /etc/passwd, .ssh/id_rsa, .aws/credentials |
| Secret exfiltration | API keys and tokens leaking through tool arguments |
| Shell injection | Reverse shells, rm -rf, encoded PowerShell commands |
| Data exfiltration | Bulk SQL dumps, git push, aws s3 cp, unauthorized file transfers |
| SSRF | Calls to metadata endpoints, localhost, and private IP ranges |
| Encoding evasion | Base64 blobs and Unicode homoglyphs used to bypass pattern detection |
| Cost abuse | Runaway agent loops burning through token budgets |
| Rug-pull attacks | Tool definitions that silently change mid-session |
Clone the repository and run the setup script.
Requirements:
The setup script automatically:
mastyf shell aliasOnce installation completes, start the proxy and dashboard:
Or simply use the alias after opening a new terminal:
The dashboard will be available at:
If the dashboard is running, verify the HTTP bridge:
Full visibility into every action your AI takes.
| Section | What you see |
|---|---|
| Protection | Block rate, top triggered rules, live threat feed |
| Activity | Every tool call with full arguments, allow or block status, timestamp |
| Policy | Live rule editor with hot-reload from YAML |
| Threat Lab | AI-suggested attack tests, reviewed and approved before anything applies |
| Cost | Token usage and cost estimates broken down per tool call |
Do not expose port 4000 publicly without enabling dashboard auth. The default local dev config has
DASHBOARD_AUTH_DISABLED=true.
Every tool call passes through three layers before it reaches your infrastructure.
Layer 1 - Pattern detection Regex-based scanning for injection, dangerous paths, leaked secrets, shell commands, and encoding tricks. Runs in microseconds with no external dependencies.
Layer 2 - Schema validation Rejects malformed payloads, oversized arguments, and JSON-RPC violations before they reach policy evaluation.
Layer 3 - Semantic review An optional local LLM (Ollama) or cloud model evaluates borderline calls that pass pattern checks. Falls back to heuristics if no model is configured.
Anything that fails is blocked. The tool never runs. Everything is logged.
Your rules live in default-policy.yaml. You own them. mastyf.ai enforces them.
Roll out safely with three enforcement modes:
| Mode | Behavior | When to use |
|---|---|---|
audit | Log everything, block nothing | First week, understand what your AI does |
warn | Log and flag, still forwards | Tuning phase before enforcement |
block | Stops violations before execution | Production |
Pre-built templates for HIPAA, PCI-DSS, GxP, and data residency are in policy-templates/.
mastyf.ai runs two coordinated swarms. The CI Swarm attacks your policy before code ships. The Runtime Swarm enforces and learns from every live tool call in production. Four feedback loops connect them so the system gets harder to bypass over time.
Canonical gates: 228/228 corpus, 0 bypasses, 100% parity
Runs on every PR and nightly. Six agents work in sequence, each one hardening what the previous found.
| Agent | What it does |
|---|---|
| Scout | SAST scan, dependency audit, config review |
| Corpus | Evaluates all 228 attack fixtures against current policy |
| Evasion | Runs 120+ bypass probes and generates novel ones using an LLM |
| Parity | Verifies Node and Python implementations produce identical decisions |
| Proxy | Live stdio MCP session tests against a running proxy instance |
| Report | Writes security-swarm/latest.json with full results and metrics |
Runs inside the production proxy on every tool call.
| Component | What it does |
|---|---|
| BlockGuard | Enforces the active policy synchronously on every call. Fail-closed. |
| InstantLearner | Tracks per-block statistics and surfaces rule suggestions in real time |
| SemanticAuditor | Optional async LLM review for calls that clear pattern checks but look suspicious |
| PatternSynthesizer | Batches suggestions from InstantLearner and SemanticAuditor into candidate rules |
| Calibrator | Labels candidates, tunes thresholds, and promotes approved rules back into BlockGuard |
| Loop | Signal | Effect |
|---|---|---|
| A | CI bypass found | Added to corpus, CI now guards against it permanently |
| B | Runtime block pattern | Synthesized into a new rule, promoted to BlockGuard |
| C | Calibrator label | Used to fine-tune SemanticAuditor thresholds |
| D | CI metrics (weekly) | Updates runtime config — keeps CI and production in sync |
The proxy supports five transports: stdio, HTTP, SSE, streamable HTTP, and WebSocket.
For enterprise deployments with Redis, Postgres, and Kubernetes see docs/ENTERPRISE_DEPLOYMENT.md.
Threat Lab watches live traffic and uses a local LLM to propose new attack test cases when it detects suspicious patterns. Nothing is applied automatically. You review and approve every suggestion in the dashboard before it becomes a rule.
Approved discoveries feed back into the CI attack corpus for ongoing regression testing.
Before installing any MCP server from npm, check its trust score at https://www.mastyf.ai/certified. Scores cover CVE exposure, typo-squat risk, maintainer signals, and known attack patterns. Free, no account required.
| Command | What it does |
|---|---|
node dist/cli.js start | Start proxy and dashboard on port 4000 |
node dist/cli.js onboard | Wrap your MCP config to route through the proxy |
node dist/cli.js doctor | Health check for DB, policy, and environment |
node dist/cli.js scan --all | Scan MCP configs for CVEs and injection risks |
pnpm test | Run the full test suite |
pnpm security-swarm:fast | Quick security regression, 5 to 15 minutes |
pnpm security-swarm:analyze | Full adversarial analysis |
| Problem | Fix |
|---|---|
| Dashboard shows no data | Proxy and dashboard must share the same MASTYF_AI_DB_PATH. Default is ~/.mastyf-ai/history.db |
dist/cli.js not found | Run pnpm build |
| AI still hitting tools directly | Run node dist/cli.js onboard --apply |
| Ollama warnings at startup | Run ollama serve or remove MASTYF_AI_LLM_PROVIDER from your environment |
| npm install fails | npm publish is not live yet. Use git clone and pnpm install |