Full text of research papers by DOI, arXiv, PMC or OpenReview id, verified against the record
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
One-click editor setup isnβt available for this listing yet β we donβt have a confirmed install command, and weβd rather show nothing than point your editor at the wrong package or host. Follow the projectβs own setup instructions, linked above.
fulltext-article-downloader is a Python package for programmatically downloading the full text of research articles from a DOI, arXiv id, PubMed Central id or OpenReview id. It chains together publisher APIs, open-access indexes and repositories in a fallback sequence, checks that what came back is really the requested article, and can be used from Python, from the command line, or by an AI agent through an MCP server or a Claude Code skill.
Video tutorial: https://youtu.be/fTtc4QWMYzE
tqdm progress bar, and file logs that record which tool succeeded or why an identifier failed.fulltext-config script.fulltext-download CLI, fulltext-mcp server for any MCP client, and a Claude Code plugin with a ready-made skill.Optional extras enable additional routes and the MCP server:
| Extra | Adds |
|---|---|
mcp | the fulltext-mcp server (fastmcp) |
tls | a Chrome TLS fingerprint fallback for hosts that reject plain HTTPS clients (curl-cffi) |
springer | the Springer Nature open-access XML route (sprynger) |
preprints | bioRxiv/medRxiv downloads through paperscraper |
aps | the APS route that reuses your browser's login cookies (browser-cookie3) |
all | everything above |
tls and aps both work by making a request look more like your own browser than a script: tls matches Chrome's TLS fingerprint for hosts that turn away plain HTTPS clients, and aps reuses the APS session cookie you are already signed in with. Neither opens anything you are not licensed for, but both go a step beyond a plain API client, so they are worth checking against your institution's agreements before you enable them. A default install has neither; all includes them.
For development, clone the repository and run pip install -e ".[dev,all]". Make sure to configure your installation afterwards (see next section).
All credentials are optional; each one unlocks a route. Without keys, and off campus, expect open-access papers, preprints and DOE-funded manuscripts to work and most paywalled articles to fail.
| Service | Environment variable | Where to get the key |
|---|---|---|
| Unpaywall and Crossref contact email (your own address) | UNPAYWALL_EMAIL | (enter your email address) |
| Elsevier API | ELSEVIER_API_KEY | https://dev.elsevier.com |
| Wiley TDM API | WILEY_API_KEY | https://onlinelibrary.wiley.com/library-info/resources/text-and-datamining |
| Springer Open Access API | SPRINGER_API_KEY | https://dev.springernature.com |
| Semantic Scholar (optional; raises the rate limit and enables the title search that finds arXiv copies of papers whose DOI record has no PDF) | SEMANTIC_SCHOLAR_API_KEY | https://www.semanticscholar.org/product/api (free, approved by email in a few days) |
UNPAYWALL_EMAIL is sent only to Unpaywall and Crossref, which ask API users for a contact address; Crossref serves requests that carry one from its faster "polite" pool. Keys and the email stay on your machine.
Set these environment variables or run the interactive helper:
The script stores keys in ~/.fulltext_keys, which are loaded automatically on import. If a required key is missing, the corresponding tool is skipped and the downloader falls back to other methods. Publisher keys return paywalled content only when the key or the network is entitled; the package detects truncated or abstract-only responses and moves on.
Any of these forms is accepted everywhere an identifier is expected:
10.1021/jacs.3c13302, https://doi.org/10.1021/jacs.3c13302, doi:10.1021/jacs.3c133022310.19377, arXiv:2310.19377, https://arxiv.org/abs/2310.19377, cond-mat/9712061, 10.48550/arXiv.2310.19377PMC6561843fNyXCCZ0g6The command prints one JSON object per identifier and exits 0 only when every download succeeded:
The original form fulltext-download <DOI> <OUTPUT_DIR> [<FILENAME>] still works.
fetch returns the same structure the CLI prints; fetch_many downloads a list concurrently:
The earlier functions are unchanged: download_article(doi, output_dir, ...) returns the path or raises, and bulk_download_articles(dois, output_dir, log_file=..., sleep=..., workers=...) returns a dict of paths or "ERROR: ..." strings with a progress bar.
Options shared by all of them: output_filename (single download), tools (a list of route names that overrides the publisher default), log_file (append a download log), and for fetch also check=False to skip verification and skip_existing=False to re-download a file that is already present.
fetch(..., supplements=True), fetch_many(..., supplements=True), fulltext-download --supplements and the MCP tools' supplements=true also fetch the article's supplementary files (supporting information, data tables, videos). They are separate files from separate places, so they are saved next to the article as <name>_si1.pdf, <name>_si2.xlsx, ... and listed in the result's supplements. Sources: the ChemRxiv and Elsevier APIs, Europe PMC's supplement bundle for PMC articles, and otherwise the article's landing page (Springer Nature, Wiley, bioRxiv, ACS, RSC, PLOS and others link them there; most publishers serve supplements without a subscription). Off by default because it costs one more request per article; a missing supplement never fails the download.
fulltext-mcp exposes two tools, get_paper(identifier, output_dir="", tools=None, supplements=False) and get_papers(identifiers, output_dir="", max_workers=4, supplements=False), returning the structure shown above. It needs the mcp extra and reads the same keys as the CLI. FULLTEXT_OUTPUT_DIR sets where files go when a call gives no output directory (default ./papers).
Claude Code:
Any other MCP client, in its server configuration:
uvx fetches the package from PyPI into an isolated environment on first use, so nothing needs to be installed beforehand. The server is also listed in the official MCP registry as io.github.computron/fulltext-article-downloader.
This repository is also a Claude Code plugin. It installs a skill that teaches Claude when and how to use fulltext-download, plus the MCP server above:
The skill alone (no MCP) is enough inside Claude Code: it runs the CLI and reads the JSON. The MCP server is for clients that cannot run shell commands, and for other agent frameworks.
Many articles are not open-access, and publishers explicitly restrict or discourage text and data mining. This example is expected to FAIL:
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/fulltext-article-downloader)<a href="https://allmcps.com/mcp/fulltext-article-downloader"><img src="https://allmcps.com/api/badge/fulltext-article-downloader?style=directory" alt="Fulltext Article Downloader on AllMCPs" /></a>