Replayable browser automation: Jev picks the element, a gate blocks risky actions, every run traces.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
One-click editor setup isnβt available for this listing yet β we donβt have a confirmed install command, and weβd rather show nothing than point your editor at the wrong package or host. Follow the projectβs own setup instructions, linked above.
Page truth for browser agents β and decisions that replay, test and audit.
https://github.com/user-attachments/assets/502550ec-77d4-439e-b334-fe7007f946b9
jevnav is a browser layer for agents and tests. It reads a page as facts
(structure, computed styles, the controls on screen), lets
Jev β TypeSafe's model for structured
questions β pick the element for an intent with a calibrated probability,
gates risky or uncertain actions to a human, and records every decision in a
trace that replay re-checks offline in CI.
outline, styles and diff return what the
browser resolved β font-size 32px β 28px between a mockup and the running
app is something an agent can fix. No screenshots in the decision loop.replay exit 1 β no model call, no API key.Why it exists β selector tests break when a label changes; LLM browser agents
are confident, unauditable and occasionally wrong: docs/why.md.
Requires Python 3.10+. Deciding (go, run, browse, goal) needs a TypeSafe
API key in TYPESAFE_API_KEY or ~/.config/typesafe/apikey.txt. replay,
diff, outline and styles need no key β that is the point.
1. Let Jev drive β state a goal and the outcome that proves it:
One Jev request per step; every step is gated and traced. The loop stops when
the goal is met, when nothing on the page can make progress (stuck), when the
gate wants a human (review), when the page stops changing (no_progress), or
at --max-steps. done is a claim β --success turns it into evidence
(verified, or unverified and the run fails). --dry-run decides without
acting.
2. Or script the flow and let Jev resolve each intent:
3. Replay it in CI β offline, deterministic, no key:
Change Sign in to Log in on the site and the same replay fails:
Exit code 1, with the reason. That is the regression test.
The facts a coding agent needs about a rendered page, without a screenshot:
outline(selector) for a region's structure (tags, headings, text, boxes),
styles(selector, props) for the computed values, page_state() for the
controls jevnav can act on.
jevnav diff compares a mockup with the running app as facts and exits 1 on
drift:
| element | property | mockup | app |
|---|---|---|---|
| h1 [Pricing] | font-size | 32px | 28px |
| button#cta [Start free] | border-radius | 8px | 4px |
The report also lists structure differences (missing, new and moved elements;
boxes compared with a 4px --tolerance). The loop for "here is a new UI, update
the codebase": the agent reads both pages with jevnav, edits the code itself,
re-runs diff until it exits 0, then pins the outcome with
goal(..., success="<selector>") so replay --execute keeps checking it.
jevnav reports; it never edits your repository and never compares pixels.
Every decision gets one of three verdicts:
| verdict | meaning |
|---|---|
auto | confidence at or above the threshold and nothing risky β the action runs |
review | a human confirms first: low p, a risky intent, or a truncated candidate list |
blocked | no decision was possible (the model answered none, or the call failed) |
review and blocked never execute. Thresholds and risky patterns live in an
optional gates.yaml; defaults ship for nine languages:
The goal loop uses a lower threshold on purpose: measured correct loop decisions
land at p 0.41β0.99 and wrong ones at 0.39β0.47, so its safety comes from
deterministic checks instead β fill on a button is refused, a field with no
context value is blocked, two steps that change nothing stop the run, risky
patterns always go to review, and the outcome is verified against --success.
Or, for Cursor, Claude Desktop, VS Code and other clients:
No URL or flags needed: the agent opens pages with goto, one server serves
every site, and each session writes an auditable jevnav-session.trace.jsonl
(--no-trace opts out). The deciding tools are what no other browser MCP has:
| tool | what it does |
|---|---|
browse(intent, action, value, min_confidence) | one step: Jev picks the element, the gate decides, only auto acts |
goal(goal, context_json, max_steps, success) | drive the whole way; returns done / stuck / review plus the verification |
goto(url) Β· page_state() Β· summary() | open a page, list what jevnav can act on, session totals |
Plus 28 acting and inspecting tools (forms, keys, uploads, tabs, console,
network, styles, outline, emulation, tracing, Lighthouse), each with MCP
annotations so the host knows which calls change state. A decision costs about
$0.00004 and ~330ms, and the page never enters the LLM's context. Full tool
reference, security flags and when to pick jevnav vs. Playwright or
chrome-devtools-mcp: docs/mcp.md.
No model call, no API key, ~30 seconds. Fails when a recorded target changed,
became ambiguous, or a recorded --success selector is no longer visible.
Inputs: trace, report, execute, json, version (default latest from
PyPI, or local for a checkout). @v0 floats; pin a release tag such as
@v0.2.2 for fully reproducible CI.
The jev fixture ships with the package: an ordinary Playwright test gets Jev
decisions, and every test writes a trace that replays in CI.
A review verdict fails the test before the action runs, ${VAR} values are
recorded by name only, and jev.page is the real Playwright page for everything
else. Runnable example with a committed trace:
examples/pytest-interop/.
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/jevnav)<a href="https://allmcps.com/mcp/jevnav"><img src="https://allmcps.com/api/badge/jevnav?style=directory" alt="Jevnav on AllMCPs" /></a>