The full upstream README, mirrored here for reference. Install config, tool schemas, adoption signals, and an original overview live on the PDF Triage listing page.
An MCP server that lets any AI tool read local PDFs — without uploading them, without an OCR bill, and without silently handing back garbage.
Built on @firecrawl/pdf-inspector (Rust, no ML models, no external services).
Most PDF tooling has the same failure mode: it returns confident text regardless of whether extraction actually worked. Broken CID fonts, substitution-cipher encodings, scanned pages with no text layer — you get plausible-looking output and find out downstream, if at all.
pdf-inspector is unusually good at knowing when it failed. It emits U+FFFD rather than guessing at an unmapped CID, runs substitution-cipher detection over its own output, and reclassifies a document as scanned when extracted text drops below 50% alphanumeric. But it stops at reporting those findings on a result object — and most wrappers throw them away.
This server acts on them. Every response carries the trust signals, above the content, where the model reads them first:
Three design rules follow:
pdf_classify costs ~20ms and tells you whether extraction is worth attempting at all.~/.ssh/id_rsa. Enforced in code, not left to the model's judgement.Nothing to install. Every config below runs the published package straight from npm:
Your MCP client runs that for you — you only need to paste the config. Confirm it works first:
Requires Node 20+. Available on npm as pdf-triage-mcp and in the MCP Registry as io.github.vishalmeena2211/pdf-triage-mcp.
Then substitute "command": "node", "args": ["/absolute/path/to/dist/index.js", ...] for the npx invocation in any config below.
Every config below is complete as written except for one value:
/Users/me/Documents — replace with the directory the server may read. This is the only thing you must change.Use an absolute path; ~ is not expanded by most clients. Repeat --root for multiple directories.
PATH gotcha, applies to every GUI client below. Desktop apps launch servers with a minimal environment, so bare
npxoften fails to resolve even though it works in your terminal. If the server won't start, substitute the absolute path — find it withwhich npx(commonly/opt/homebrew/bin/npxon Apple Silicon,/usr/local/bin/npxon Intel macOS).
Why
-y? It skips npx's install confirmation prompt. Without it, a first run can hang waiting for input that an MCP client cannot provide — the server appears to start and then silently times out.
Config file
| OS | Path |
|---|---|
| macOS | ~/Library/Application Support/Claude/claude_desktop_config.json |
| Windows | %APPDATA%\Claude\claude_desktop_config.json |
| Linux | ~/.config/Claude/claude_desktop_config.json |
Verify: Fully quit and relaunch Claude Desktop (not just close the window). A tools icon appears near the chat input — click it and confirm the four pdf_* tools are listed.
Logs: ~/Library/Logs/Claude/mcp*.log (macOS), %APPDATA%\Claude\logs\mcp*.log (Windows).
CLI — the easiest route. The -- separator is mandatory; everything after it is the server command.
Or edit .mcp.json at the project root directly:
| Scope | Stored in | Shared |
|---|---|---|
local (default) | ~/.claude.json, under this project | No |
user | ~/.claude.json, top level | No |
project | .mcp.json in repo root | Yes, via git |
Verify: claude mcp list → look for ✔ Connected. Project-scoped servers need approval on first use — run /mcp inside a session.
Config file: .cursor/mcp.json (project) or ~/.cursor/mcp.json (global). Project wins on conflict.
Verify: Cursor hot-reloads — no restart. Open Cursor Settings → Tools & MCP and look for a green dot next to pdf-triage.
Config file: ~/.codeium/windsurf/mcp_config.json (macOS/Linux), %USERPROFILE%\.codeium\windsurf\mcp_config.json (Windows).
Not created on first launch — create it yourself if missing.
Verify: Windsurf watches the file and hot-reloads on save. Tools appear in Cascade on the next chat session.
The key is
servers, notmcpServers. This is the most common mistake when copying a config from Claude Desktop.
Config file: .vscode/mcp.json (workspace), or Command Palette → MCP: Open User Configuration (global).
CLI alternative:
Verify: MCP tools only work in Agent mode — switch from Ask/Edit to Agent in Copilot Chat, then click Configure Tools and confirm the pdf_* tools appear. Restart VS Code after first adding the file.
The key is
context_servers, notmcpServers, andcommandis a nested object rather than a string.
Config file: ~/.config/zed/settings.json (macOS/Linux), %APPDATA%\Zed\settings.json (Windows). Command Palette → zed: open settings.
If your Zed version rejects that, it predates the nested form — try command, args and env flat at the top level of the server object instead.
Verify: Agent Panel (Cmd+Shift+A) → gear icon → MCP Servers. Green dot means connected.
Config file — separate from VS Code's own:
| OS | Path |
|---|---|
| macOS | ~/Library/Application Support/Code/User/globalStorage/saoudrizwan.claude-dev/settings/cline_mcp_settings.json |
| Windows | %APPDATA%\Code\User\globalStorage\saoudrizwan.claude-dev\settings\cline_mcp_settings.json |
| Linux | ~/.config/Code/User/globalStorage/saoudrizwan.claude-dev/settings/cline_mcp_settings.json |
autoApprove runs the listed read-only tools without a confirmation prompt.
Easier route: Cline panel → MCP servers icon → Edit MCP Settings opens this file directly.
Verify: Panel refreshes automatically; green dot next to the server.
Config file: ~/.continue/config.yaml (global) or .continue/config.yaml (project). YAML is current; config.json is deprecated.
mcpServershere is a list, not an object — and YAML needs spaces, never tabs.
Verify: Reloads automatically on save. Switch Continue to Agent mode — MCP tools are unavailable in other modes.
CLI:
Or edit ~/.gemini/settings.json (global) / .gemini/settings.json (project):
Verify: Run /mcp inside a gemini session — servers show CONNECTED with their tool list. Or gemini mcp list from the shell.
Config file: ~/.codex/config.toml (global) or .codex/config.toml (project).
TOML, and the key is
mcp_servers— snake_case, nevermcpServers.
Verify: codex doctor --json validates the config syntax. Note that it validates syntax only — it does not confirm the server actually spawned.
Known upstream issue: several Codex CLI versions have a bug where stdio servers validate cleanly but silently fail to start, showing
Tools: nonein the TUI (#3441, #26810). That is a Codex runtime bug, not a config error.
AI Assistant — configured through the IDE, no file to edit:
Junie uses a file instead — ~/.junie/mcp/mcp.json (global) or .junie/mcp/mcp.json (project), same JSON shape.
Verify: Check the Status column in the MCP settings panel; click it to list the server's tools.
Config file: ~/.lmstudio/mcp.json (macOS/Linux), %USERPROFILE%\.lmstudio\mcp.json (Windows).
Easier via the app: right sidebar → Program tab → Install → Edit mcp.json.
Verify: Auto-reloads on save; tools appear in the Program panel. LM Studio shows a confirmation dialog the first time a model calls a tool.
| Client | File | Top-level key | Restart? |
|---|---|---|---|
| Claude Desktop | claude_desktop_config.json | mcpServers | Full quit |
| Claude Code | .mcp.json / CLI | mcpServers | No |
| Cursor | .cursor/mcp.json | mcpServers | No |
| Windsurf | ~/.codeium/windsurf/mcp_config.json | mcpServers | No |
| VS Code Copilot | .vscode/mcp.json | servers | First time |
| Zed | ~/.config/zed/settings.json | context_servers | No |
| Cline | cline_mcp_settings.json | mcpServers | No |
| Continue.dev | ~/.continue/config.yaml | mcpServers (list) | No |
| Gemini CLI | ~/.gemini/settings.json | mcpServers | No |
| Codex CLI | ~/.codex/config.toml | [mcp_servers.*] | N/A |
| JetBrains | IDE settings UI | mcpServers | No |
| LM Studio | ~/.lmstudio/mcp.json | mcpServers | No |
The three that differ: VS Code (servers), Zed (context_servers + nested command), Codex (TOML mcp_servers). Everything else takes the Claude Desktop format verbatim.
| Tool | Cost | Purpose |
|---|---|---|
pdf_classify | ~20ms | Type, confidence, page count, exact pages needing OCR. Call this first. |
pdf_extract | ~150ms | PDF → Markdown. Truncates by default; slice with pages. |
pdf_search | ~150ms | Locate text, return page-attributed snippets. Cheapest way into a long document. |
pdf_tables | ~150ms | Tables only, as Markdown pipe tables. |
Full parameter reference: docs/TOOLS.md.
The intended flow on an unfamiliar document:
Environment equivalents: PDF_TRIAGE_ROOTS (separated by the platform PATH delimiter — : on macOS/Linux, ; on Windows), PDF_TRIAGE_MAX_CHARS, PDF_TRIAGE_MAX_FILE_BYTES, PDF_TRIAGE_LOG_LEVEL. Flags win over environment.
Roots are a security boundary, not a convenience. Grant the narrowest directory that works. Paths are resolved through symlinks before checking, so a link inside a root pointing outside it is rejected rather than followed.
Upstream ships prebuilt native binaries for exactly three targets: linux-x64-gnu, darwin-arm64, win32-x64-msvc. No musl build, no Linux ARM64 build (upstream #216) — so it fails to load on Alpine containers, Graviton instances, and most edge runtimes.
This server prefers native and falls back to WASM, which runs anywhere. Capability differences are surfaced, never faked:
| Native | WASM | |
|---|---|---|
| Classify / extract | Yes | Yes |
| Per-page extraction | Yes | No — throws, and pdf_search reports its matches are unattributed |
pages selection | Yes | No — ignored, and the response says so |
Check which engine you got: pdf_classify reports it, and the server logs engine selected at startup.
Inherited from upstream. Worth reading before you trust output:
text_based with high confidence (#212). This server detects and escalates it — the one upstream failure mode we actively guard.pypdf and pdfium silently repair will throw (#228).The TypeScript config runs every strictness flag including exactOptionalPropertyTypes and noUncheckedIndexedAccess. Upstream responses are validated with Zod at the boundary rather than cast — see docs/ARCHITECTURE.md for why.
Debug a client connection:
Troubleshooting: docs/TROUBLESHOOTING.md.
pdf_regions — bbox-scoped extraction for hybrid model pipelines{page, bbox} for visual citation UXMIT