Skip to main content
AllMCPs
BrowseBestCategoriesStackCompareToolsGuidesBlog
Log in Submit MCP

Stay in the loop

Get new MCP servers and top picks in your inbox.

AllMCPs

The open directory for discovering and installing Model Context Protocol servers.

AllMCPs on GitHub (opens in a new tab)
Launched onTiny Startupstinystartups.com
Explore
  • Browse servers
  • Best MCP servers
  • Categories
  • MCP clients
  • Agent prompts
  • Stack Builder
  • Compare servers
  • Random discovery New
  • Submit a server
  • Pricing & Boost Boost
Learn
  • Guides hub
  • What is MCP?
  • Install guide
  • Build an MCP server
  • Deploy an MCP server
  • Security guide
  • Troubleshooting
  • MCP for SEO & AEO
  • Protocol versioning
  • Transports: stdio vs HTTP
  • State of MCP (stats)
  • Blog & updates
Tools
  • All developer tools
  • Config generator
  • Config validator
  • Config auditor
  • MCP playground
  • Token calculator
  • OpenAPI β†’ MCP
  • Badge generator
For agents
  • REST API docs
  • Trust & traffic Live
  • Remote MCP server SSE β†— (opens in a new tab)
  • llms.txt β†— (opens in a new tab)
  • Catalog JSON β†— (opens in a new tab)
Company
  • About
  • Advertise Sponsor
  • Contact
  • GitHub β†— (opens in a new tab)
  • Terms
  • Privacy
AllMCPs VerifiedAllMCPs VerifiedFeatured on Nick LaunchesFeatured on Nick LaunchesLaunch Llama NewsletterLaunch Llama NewsletterVerified DR - allmcps.comVerified DR - allmcps.comFeatured on SaaSGrowFeatured on SaaSGrowFeatured on Twelve ToolsFeatured on Twelve ToolsFeatured on Saaspa.geFeatured on Saaspa.geFeatured on Findly.toolsFeatured on Findly.toolsFeatured on Startup FameFeatured on Startup FameFeatured on LaunchKiwiFeatured on LaunchKiwiFeatured on ScrollLaunchFeatured on ScrollLaunchFeatured on DailyPingsFeatured on DailyPingsFazier badgeFazier badgeFeatured on NewTool.siteFeatured on NewTool.siteFeatured on saasfame.comFeatured on saasfame.comDR Checker - Domain RatingDR Checker - Domain RatingListed on Turbo0Listed on Turbo0Launched on LaunchBoard - Product Launch PlatformLaunched on LaunchBoard - Product Launch PlatformList on SimilarlabsList on Similarlabshttps://codetrendy.comhttps://codetrendy.comListed on DevTool.ioFeatured on BuildlistFeatured on BuildlistLaunched on Tiny StartupsFeatured on ShowMeBestAIFeatured on ShowMeBestAIFind us on LaunchZoneFind us on LaunchZoneAllMCPs VerifiedAllMCPs VerifiedFeatured on Nick LaunchesFeatured on Nick LaunchesLaunch Llama NewsletterLaunch Llama NewsletterVerified DR - allmcps.comVerified DR - allmcps.comFeatured on SaaSGrowFeatured on SaaSGrowFeatured on Twelve ToolsFeatured on Twelve ToolsFeatured on Saaspa.geFeatured on Saaspa.geFeatured on Findly.toolsFeatured on Findly.toolsFeatured on Startup FameFeatured on Startup FameFeatured on LaunchKiwiFeatured on LaunchKiwiFeatured on ScrollLaunchFeatured on ScrollLaunchFeatured on DailyPingsFeatured on DailyPingsFazier badgeFazier badgeFeatured on NewTool.siteFeatured on NewTool.siteFeatured on saasfame.comFeatured on saasfame.comDR Checker - Domain RatingDR Checker - Domain RatingListed on Turbo0Listed on Turbo0Launched on LaunchBoard - Product Launch PlatformLaunched on LaunchBoard - Product Launch PlatformList on SimilarlabsList on Similarlabshttps://codetrendy.comhttps://codetrendy.comListed on DevTool.ioFeatured on BuildlistFeatured on BuildlistLaunched on Tiny StartupsFeatured on ShowMeBestAIFeatured on ShowMeBestAIFind us on LaunchZoneFind us on LaunchZone
Β© 2026 Jackalope Digital LLC. All rights reserved.
  1. Home
  2. πŸ’» Developer Tools
  3. Sidecrew
S
Health: Not checked yetWe have not completed a health check for this listing yet.No health check has run yet.

Sidecrew

User RatingsBe the first to rate and review this MCP server! Enrichment pendingWe haven’t run our AI enrichment pass on this listing yet, so the overview, use cases, and FAQ below may be sparse or missing. We work through the catalog over time β€” check back soon.
View Repository

Code changes and unit tests by local MLX models, behind a gate a machine can run.

Quick Install

Automated & IDE Setup

Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β€” or use 1-click editor setup below.

One-click editor setup isn’t available for this listing yet β€” we don’t have a confirmed install command, and we’d rather show nothing than point your editor at the wrong package or host. Follow the project’s own setup instructions, linked above.

Manual Client & Custom JSON ConfigExpand JSON β–Ύ
No confirmed setup config for this listing yet. We only publish a config block when the install details come from the project itself β€” its README, its docs, or a verified owner. We haven’t found those for sidecrew, and we’d rather show nothing than a guess you’d paste into your client. Follow the project’s own setup instructions for the current steps.
Install Directory Badge Claim listing AlternativesπŸ’» More in Developer Tools

Documentation Overview

sidecrew

You ask for a change to your source code, and you get it β€” made on your Mac, by models that cost you no Claude tokens, and only after a machine has checked it.

Claude Opus plans the work and reviews what came back. Small models running locally (MLX, no network) do the narrow parts. Between them sits a gate a machine can run β€” for a code change: the diff stayed where it was asked to, tsc is clean, and every test that passed before still passes. Claude never sees anything that did not survive it.

console
$ sidecrew fix change_plan.json
fix 2026-09-20T09-14-02Z-rename: 1 step, 12 tasks, 13.0 GB free Γ· 6.5 GB per slot = 2; 1 worker up
step 1/1 "rename the helper": capturing the baseline
rename-01: survived (src/util/asset.util.ts)
rename-02: survived (src/util/asset.util.ts, src/util/asset.util.spec.ts)
…
multi-06: failed at compile Β· escalating β€” renamed the symbol 7 times over, 5 compile errors
11/12 survived Β· 0 β†’ 0 tsc errors Β· .sidecrew/runs/2026-09-20T09-14-02Z-rename/result.json

Illustrative shape, not a captured transcript β€” it is the real output format with the counts from the React measurement below, and no client's file names. 11 of 12 with a 95 % interval of [0.615, 0.998] is the number; the line above is what it looks like while it runs.

Unit tests are workload #1 β€” the same shape with a mutation-kill gate instead β€” and the numbers for both are below, each with its interval.

sidecrew needs 24 GB of installed RAM. Below that a 7B cannot sit beside a normal working set (measured), so sidecrew refuses to run and tells you why rather than quietly becoming a different tool (ADR-0073). sidecrew doctor answers this before you plan anything.

And the other half of the ledger, measured 19–20 Sep 2026: the workers are free, the coordination is not. Opus planning costs 8,400–22,100 tokens per task, depending almost entirely on how many tasks are in the plan β€” see What this costs to run below. A page that says "the work is free" without saying that is telling you half of it.

Before any number below: how tightly to read it

Every rate on this page is an interval, never a point, and the reason is measured. Replaying a frontier model's own diffs β€” which the gate finds no defect in β€” through the gate a second time, it disagreed with itself on 2 of 19 candidates: D = 0.105, 95 % interval [0.013, 0.331] (experiments/gate-error-rate/, against a rule frozen before the first replay). That rule's own verdict at D > 0.10 is UNACCEPTABLE, and its consequence is this paragraph: no survival rate here is published as a bare fraction, D travels beside each of them, and no two of our rates are compared without asking whether their difference survives it.

What that does and does not damage. D counts the gate refusing changes that were fine. Nothing in it is evidence of the gate admitting something bad. So:

The gate admits nothing that fails β€” stands. The gate refuses only things that fail β€” measured false, at roughly one evaluation in ten, and unproven at any tighter bound.

Two of the three known causes are now recorded on every verdict β€” memory pressure (ADR-0066) and a baseline captured on a different calendar day (ADR-0069). The third is undiagnosed, and both measured disagreements happened on a quiet machine, so it is neither of the first two. D was measured on workload #2a's gate; it says nothing about workload #1's mutation gate, which is a different oracle (Β§4.4 of the rule).

Workload #2a β€” behaviour-preserving changes, measured

Against a decision rule frozen before the tool had produced a single verdict (experiments/go-no-go-2a/results/REPORT.md). S is survival, A is a blind reviewer's approval of the diff. Every cell is k/n with its exact 95 % interval; D = 0.105 [0.013, 0.331] applies to every S.

inputS, local 7BS, Haiku controlA, local 7BA, control
fixture3/3 [0.292, 1.000]3/3 [0.292, 1.000]1/3 [0.008, 0.906]3/3 [0.292, 1.000]
a commercial Nest/jest codebase12/12 [0.735, 1.000]12/12 [0.735, 1.000]9/10 [0.555, 0.997]10/10 [0.692, 1.000]
a commercial React/jest codebase11/12 [0.615, 0.998]12/12 [0.735, 1.000]9/10 [0.555, 0.997]10/10 [0.692, 1.000]

Read the intervals, not the fractions: at these sample sizes 12/12 and 11/12 are not distinguishable, and neither is the 7B from the control on A. Both real projects passed the frozen rule by a margin of exactly zero, on a sample of ten. The survival rate decided nothing β€” five of six measurable cells sit at or above 0.917, so the rule's ratio clauses were inert, and the only thing that separated a local 7B from a network model was the approval rate.

What that approval rate is about, measured: the local model makes unrequested cosmetic edits β€” a deleted docblock, a reworded comment, a stray blank line β€” in 4 of 23 [0.050, 0.388] sampled survivors, against the control's 0 of 23 [0.000, 0.148]. confined ∧ compiles ∧ tests pass cannot see any of it, which is the point.

What you gain over just asking Opus

Measured 18–19 Sep 2026 against the thing a colleague actually does today β€” asking a frontier model directly, one task at a time, with no plan and no gate. Same 19 hand-written tasks, same two codebases, four arms.

Opus aloneSonnet alonesidecrew
tokens, 19 tasks, Nest codebase80,13188,8200
tokens, 19 tasks, React codebase95,33588,6800
output vs Opus, Nest codebaseβ€”identical 19/19identical on 18 of 19
changes delivered and verified, Nest19 (unverified)19 (unverified)19/19 [0.824, 1.000] survived
changes delivered and verified, React19 (unverified)19 (unverified)15/19 [0.544, 0.939] survived; 4 escalated

The token columns are counts and are exact. The survival rows are rates, so they carry their intervals, and D = 0.105 [0.013, 0.331] applies to both of them.

A fourth arm ran the same gate over Opus's own diffs, to separate the gate's contribution from the worker's: 19/19 [0.824, 1.000] on the Nest codebase and 18/19 [0.740, 0.999] on the React one. Two things follow, and the second is the more useful.

The gate found nothing wrong with a frontier model's work. Its single rejection there is a tsc fragility it rejects from both arms β€” so on these tasks the gate adds no safety on top of Opus. Its entire value is that it makes the free worker usable, not that it second-guesses the expensive one.

And it is what separates a worker defect from a project defect. Of the four React failures, three vanish when Opus writes the diff β€” genuine worker defects β€” and one reproduces exactly, because the 7B's output for that task was byte-identical to Opus's. Worker-attributable survival is therefore 15/18 [0.586, 0.964], not 15/19. Without that control the small model would have been blamed for a third more failures than it caused β€” and note that the two intervals overlap almost entirely, so the correction changes the attribution rather than the number.

Three things that buys you, and one it costs.

  1. The work is free. Not cheaper β€” zero worker tokens. Opus plans and reviews survivors; the editing happens on your Mac.
  2. On work a compiler and a suite can check, you give up nothing for it. Eighteen of nineteen diffs were byte-identical to what Opus produced. Not "comparable quality" β€” the same file.
  3. When it is wrong, you get told, loudly, and nothing is applied. The four React failures were a symbol renamed seven times over (5 compile errors), a rename that also hit a type, a module path and a public property key (122 compile errors), a change that renamed the file it was given, and one that a frontier model produced identically. Each printed as escalated, each with the exact errors, none merged. Asking Opus directly gives you a diff and the sentence "nothing else changed" β€” accurate here, and unverified.
  4. It costs wall-clock. ~25 s per task interactively with Opus; 216–271 s with sidecrew (15 s to generate, 200–255 s for the project's own suite and tsc). It is free and unattended, not fast.

What is not claimed. The frozen rule's verdict is WITHHELD: every task in these runs is a rename, and the rule refuses a verdict until at least half are null guards, API migrations or dead-code removal. The measurement vindicated that clause rather than surviving it β€” two frontier models and a 7B emitted identical bytes on 18 of 19 Nest tasks, so that task set cannot rank anything. experiments/status-quo/ has the arms, the protocol frozen before they ran, and the defects.

Requirements

OSmacOS on Apple silicon. The local worker is MLX, which is Apple's. Not a port away β€” a different inference stack.
RAM24 GB installed or more for the local tier and everything this page measures. The threshold is on installed RAM, never free, so a machine that can host a worker never falls back silently.
Nodeβ‰₯ 20. A project whose own engines demands newer is honoured β€” run sidecrew under the project's Node (ADR-0049 cost a night to learn).
Pythonmlx-lm, installed by you: pip install mlx-lm. Deliberately not bundled β€” external capabilities are shelled out and reported by doctor, never vendored.
Disk~4 GB for the 7B (downloaded on first sidecrew serve, pinned revision), plus ~500 MB per verification sandbox while a run is in flight.
The projectTypeScript with a green-ish tsc and a test suite that runs. Workload #1 also does Swift/XCTest.

Read the full README β†’View source on GitHub β†’

Related MCP Servers

View all in Developer Tools View all alternatives
  • O
    Openapi MCP Server

    Connect any HTTP/REST API server using an Open API spec (v3)

    πŸ’» Developer Tools3 views
    Compare vs Openapi MCP Server β†’
  • C
    Claude Task Master

    AI-powered task management system for AI-driven development. Features PRD parsing, task expansion, multi-provider support (Claude, OpenAI, Gemini, Perplexity, xAI), and selective tool loading for optimized context usage.

    πŸ’» Developer Tools8 views
    Compare vs Claude Task Master β†’
  • M
    MCP Server Docker

    Integrate with Docker to manage containers, images, volumes, and networks.

    πŸ’» Developer Tools3 views
    Compare vs MCP Server Docker β†’
  • V
    Vibe Hnindex

    Index source code into a local knowledge base, search with keyword + semantic + hybrid modes.

    πŸ’» Developer Tools2 views
    Compare vs Vibe Hnindex β†’

Reviews

No reviews yet β€” be the first to share how this listing worked for you.

Frequently Asked Questions about Sidecrew

We don't have a confirmed install command for sidecrew yet, so we don't publish a generated one β€” a guessed package name would point at the wrong package or none at all. Follow the project's own README or setup instructions (https://github.com/lvlrSajjad/sidecrew) for the current steps.

AllMCPs Directory Badge

Full Badge Customizer

Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.

Badge Style:
Live Dynamic SVG PreviewSidecrew AllMCPs Directory Badge
Markdown (GitHub README)
[![AllMCPs](https://allmcps.com/api/badge/sidecrew?style=directory)](https://allmcps.com/mcp/sidecrew)
HTML Embed
<a href="https://allmcps.com/mcp/sidecrew"><img src="https://allmcps.com/api/badge/sidecrew?style=directory" alt="Sidecrew on AllMCPs" /></a>

Technical Specs & Signals

CategoryπŸ’»Developer Tools
More technical detailsExpand β–Ύ
Last updatedSep 28, 2026
Views0
Unique ViewsTotal visits recorded for this listing page on AllMCPs.
Installs0
Installs & Copy ActionsTotal times users copied install commands or configuration snippets for this server.
27Quality signal: Emerging Β· 27/100How this signal is calculated β–Ύ
Server availabilityNot measured

Not scored for repo-hosted servers β€” we can't reach the running server, only its GitHub page. Hosted MCP endpoints are health-checked live.

Verified ownership8/20
Documentation & tools11/30
Adoption & activity1/15
Community engagement0/10

A guidance signal from public completeness & health data β€” not a user rating. New listings start lower and rise as they add docs, get verified, and grow adoption. Signals we can't observe for a listing are skipped, not counted against it.

β˜… Spotlight Slot

Feature Your MCP Server

Get maximum visibility for your server across our directory, search results, and detail pages.

Spotlight Your Server

Own this project?

This directory is pre-filled from public sources. Claim via GitHub README, site badge, or DNS TXT to unlock edit access and the Official badge and attach your website β€” proof is checked automatically, then reviewed by our team.

Free dofollow backlink: add your website and place the AllMCPs badge on it β€” no claim needed. We detect it automatically and keep it verified as long as the badge stays live.

Claim & get free dofollow

Share & Embed

Add our SVG badge (dark/light directory styles) or embeddable widget to your site.

Explore more

More in πŸ’» Developer Tools β†’Best MCP servers for Developers β†’Alternatives to Sidecrew β†’Install in Claude DesktopInstall in CursorInstall in VS CodeSetup guides for all 13 MCP clients