Produce verifiable FORRT nanopublication chains: grounded quote checks and Science Live projections.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
One-click editor setup isnβt available for this listing yet β we donβt have a confirmed install command, and weβd rather show nothing than point your editor at the wrong package or host. Follow the projectβs own setup instructions, linked above.
An MCP server for producing verifiable FORRT nanopublication chains β the tools a researcher needs while doing the work, not while looking for it. It serves all three shapes of study on the same rails: reproduction, replication, and research that starts from scratch.
It is the third of three servers that compose:
| Server | Question it answers |
|---|---|
| OpenAIRE MCP | What is in the literature? |
replication-radar | What is worth replicating, and has it been done? |
forrt-research-mcp | How do I produce a chain that is correct and verifiable? |
Built for the workflow in
forrt-replication-template,
but it needs nothing from that repo β any agent can call it.
All three work here, and they use the same chain:
Step 01 is one of three anchors, and they are not variants of one form. Which anchor fits is independent of whether the study is a reproduction, a replication or original research:
| Anchor | Required fields | Notes |
|---|---|---|
01_quote Quote-with-comment | paper, quotation, comment | no type, no label. quotation is capped at 500 characters and comment at 800 β in the template's own regex, not in prose |
01_pico PICO question | label, description, type, + P/I/C/O descriptions | the only anchor with a question-type vocabulary (5 terms) |
01_pcc PCC question | label, description, + P/C/C descriptions | no type field at all |
The three share no field ids whatsoever, so call template_fields on the
anchor you are actually using rather than assuming they match. The
pico_question_type vocabulary is named for PICO deliberately: a PCC question
has no type to look up.
verify_quote applies to the Quote anchor only β and that is the one anchor
with a dedicated tool, because a quotation is the one field whose correctness
can be proved rather than reviewed.
constellation and prior_work surface a question anchor with its framework,
label, full question text and its framework's components β PICO's
population / intervention / comparator / outcome, PCC's population / concept /
context. Question nodes keep all of that in a question object and leave their
top-level label empty, so anything reading them like a quote returns nothing.
Tested against two real published question nanopubs.
For a study starting from scratch, the Claim is your own hypothesis, derived
from your own question or from the work you are building on. The FORRT Claim
template's source is optional, so it needs no external paper β and the
Claim-before-Study order then reads as
pre-registration, not as a mismatch. verify_chain takes a mode
(auto / replication / reproduction / new_research) that changes only what
is required: from-scratch research has no existing work to cite, so no CiTO
step and no cited DOI are expected.
The one place the templates still assume an original is the study_type
vocabulary on 04_study, whose three terms are all replication-flavoured
(Replication Study, Reproduction Study, or both). There is no term for an
original study testing its own claim. Three field prompts read oddly too β
scope and methodology say "is reproduced/replicated", and the Outcome's
conclusion says "about the original claim" β but they are only wording;
validationStatus (validated / partially supported / contradicted /
inconclusive / not tested) describes testing your own hypothesis perfectly well.
So closing the gap is plausibly one added vocabulary term and three reworded
prompts, not a new template family. When that lands, this server picks it up
with no code change: template_fields and vocabulary fetch live, and
driftedFromSnapshot flags the supersession so the vendored copy gets re-cut.
A note on using "Reproduction Study" for from-scratch work. It is a reasonable workaround while the vocabulary lacks a better term, but the template means the replication-science sense β "direct reproduction: same methodology, same tools", i.e. re-running someone else's analysis β not the RSE sense of "my work is reproducible". Downstream consumers read it the first way:
replication-radar's verdict overlay andverified_claimswill present the study as verification of an existing claim.
Two jobs here look like reasoning but are not, and doing them in an agent loop makes them unreproducible and expensive:
1. Reading a published chain. /np/constellation walks the FORRT citation
graph bidirectionally and returns everything it reaches. For a single
marine-heatwave chain that is 330 KB, 98 nodes and 1012 edges β of which 64
nodes are AIDA statements belonging to entirely different studies. Handing
that to an agent burns its context and invites it to reason over another paper's
claims. constellation returns the same chain in ~16 KB (4.7 %).
2. Checking a quotation is real. A FORRT Quote nanopub must be verbatim. That is a string search, not a judgement β so it should be a tool that cannot be talked out of its answer, and that anyone can re-run to get the same result.
It does not search papers (that is the OpenAIRE MCP), rank replication targets
(that is replication-radar), or write nanopub content. It also does not
extract claims: choosing which sentence carries a paper's headline claim is a
judgement, and wrapping a judgement in a tool would not make it reproducible β a
model behind a tool boundary is exactly as non-deterministic as one in the agent
loop. The division this server is built around:
| The model proposes | This server disposes |
|---|---|
| Which sentence is the claim | Whether that sentence is in the PDF, where, under what SHA-256 |
| How to phrase the AIDA | What fields the template actually has, and their caps |
| Which claim type or CiTO relation applies | Which values the form will actually accept |
| Which Wikidata topic is meant | Whether that QID exists and is of the right type |
| Which paper to cite | Whether that DOI resolves, and to what |
| Whether to extend or dispute prior work | What prior chains already claimed |
User-scoped (-s user) so it is available in every session and folder, and does
not clash with a per-repo config. For other agents, the stdio command is
forrt-research-mcp.
The default base is production. To test against the dev deployment, set
SCIENCELIVE_API_BASE=https://api-dev.sciencelive4all.org. A constellation that nobody has requested recently can take ~60 s to build; after that it is cached for 4 h and returns in well under a second.
verify_quote(pdf_path, quotation)Proves a candidate quotation is in a source PDF before it is published as verbatim. Returns a graded verdict with page, character offsets and the file's SHA-256:
| Verdict | Meaning |
|---|---|
exact | byte-identical to the extracted page text |
normalized | matched after whitespace / ligature / typographic punctuation / line-break-hyphen repair |
extraction_tolerant | additionally ignored hyphens and punctuation spacing |
whitespace_insensitive | additionally deleted every space β for text layers that break words apart |
not_found | not in this PDF β do not publish it |
Every tier canonicalises formatting only β never words, digits or order. A
quotation with one digit changed scores ~0.91 similarity against the source and
is still not_found; on a miss, closest.text_in_pdf shows what the paper
actually says there.
A tier below exact is normal. PDF extraction inserts line breaks, loses
hyphens, and sometimes breaks words apart entirely:
extraction_tolerant, because pypdf reads the paper's 35-year as 35year
and (p < as ( p <;CP en semble, largest c hanges, prec ipitation β spaces inside words, which no punctuation rule
can undo. That needs whitespace_insensitive, which deletes every space and
therefore does not verify word boundaries; it declines to run below 40
characters, where two short texts could collide."Character-for-character" is not literally achievable against extracted PDF text, which is why the result is graded rather than boolean.
Words split across a line break (convection-\npermitting) are rejoined at the
normalized tier. A suspended hyphen (15- and 30-minute) is deliberately
NOT rejoined β the rule requires a line break, because joining those would
corrupt the text.
constellation(uri, depth=5, max_nodes=80)The FORRT chain(s) reachable from a published nanopub URI, projected to agent-size: chains with their steps in chain order, the apex CiTO, any Research Synthesis, and the Quote/Question anchors attributable to this paper.
Read citedPaper, not the API's top-level paperDoi β see Upstream
quirks below.
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/forrt-research)<a href="https://allmcps.com/mcp/forrt-research"><img src="https://allmcps.com/api/badge/forrt-research?style=directory" alt="FORRT Research on AllMCPs" /></a>