Fetch any public web page through managed proxies, with optional JS rendering and extraction rules.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
One-click editor setup isnβt available for this listing yet β we donβt have a confirmed install command, and weβd rather show nothing than point your editor at the wrong package or host. Follow the projectβs own setup instructions, linked above.
A hosted Model Context Protocol (MCP) server that gives Claude, Cursor, Windsurf and any other MCP client one read-only tool for fetching any public web page. It goes out through managed proxies, renders JavaScript when a page needs it, and returns clean markdown, plain text, raw HTML or structured JSON, with nothing to host and no browser in your stack.
This is the fallback for sites with no dedicated API. When a site does have one in the HasData catalogue, that tool returns parsed fields and this one returns a page.
1,000 free credits every month, no card required. A plain fetch costs 1 credit, so the free tier covers 1,000 of them.
An MCP client and a HasData API key from the dashboard, free to create with no card. This is a remote server, so the simplest path is a URL and an x-api-key header, with no container to run. A client that only speaks stdio reaches it through a thin launcher, published as @hasdata/web-scraping-mcp on npm and hasdata-web-scraping-mcp on PyPI, shown below.
The server URL is the same for every client. We run it hands-on in Claude Code and Claude Desktop. The other blocks follow each client's own documented format for a remote server.
| Field | Value |
|---|---|
| URL | https://mcp.hasdata.com/mcp?apis=web_scraping |
| Transport | HTTP, streamable |
| Auth header | x-api-key: HASDATA_API_KEY |
Clients with OAuth support can add the same URL as a connector and sign in without putting a key in a config file.
Settings, then Connectors, then Add custom connector, then paste https://mcp.hasdata.com/mcp?apis=web_scraping and sign in.
For the config-file route, Claude Desktop loads only local (stdio) servers, so it reaches a remote server through a stdio launcher. The @hasdata/web-scraping-mcp package is that launcher, and it reads the key from the environment. Add this to claude_desktop_config.json:
For Python instead of Node, swap the launcher for the PyPI package, which uvx runs without a manual install:
~/.cursor/mcp.json for every project, or .cursor/mcp.json for one:
~/.codeium/windsurf/mcp_config.json. Windsurf calls the field serverUrl, not url:
.vscode/mcp.json in the workspace:
One call answers each of these. What changes between them is how much of the browser you asked for, and that is what the call costs.
| Tool | What it returns |
|---|---|
hasdata_web_scraping_web_scraping_scrapeWebPage | HTML, text, markdown, and/or JSON along with status code, extracted emails and links, CSS-selector extractions, and AI-structured fields per schema. 1 credit a call for a plain fetch, 10 with JS rendering |
One tool. The cost depends on what you turn on, and the table is in Pricing.
hasdata_web_scraping_web_scraping_scrapeWebPage
Fetch one URL.
| Parameter | Type | Required | Notes |
|---|---|---|---|
url | string | yes | The page to fetch |
outputFormat | array | Any of markdown, text, html, json. See Output formats | |
jsRendering | boolean | Render the page in a browser. On by default, and the main cost lever | |
proxyType | string | datacenter or residential | |
proxyCountry | string | US, UK, DE, IE, FR, IT, SE, BR, CA, JP, SG, IN or ID | |
headers | object | Custom request headers | |
wait | number | Milliseconds to wait after load | |
waitFor | string | CSS selector to wait for before reading | |
jsScenario | array | Actions to run on the page. See below | |
extractRules | object | CSS selectors to pull named fields | |
aiExtractRules | object | A typed schema an LLM fills from the page | |
extractLinks | boolean | Collect the page's links | |
extractEmails | boolean | Collect email addresses on the page | |
screenshot | boolean | Capture the rendered page | |
blockResources | boolean | Skip images and stylesheets | |
blockAds | boolean | Skip ad requests | |
blockUrls | array | Skip these URLs | |
includeOnlyTags | array | Keep only elements matching these selectors | |
excludeTags | array | Drop elements matching these selectors | |
removeBase64Images | boolean | Strip inline base64 images from the output |
extractRules maps a field name to a CSS selector, with @attr to read an attribute rather than text.
jsScenario is an array of actions run in order, covering click, wait, waitFor, waitForAndClick, scrollX, scrollY, fill and evaluate for arbitrary JavaScript. It needs jsRendering on.
aiExtractRules describes the shape you want and lets a model fill it from the HTML. Each key is an output field, typed as string, number, boolean, list or item for a nested object.
This is the part worth reading before your first call, because the response shape moves with outputFormat.
Ask for exactly one of markdown, text or html, and the content arrives as a plain string in text at the top level.
Include json, alone or alongside another format, and everything moves inside json, the top-level text becomes null, and the requested formats become keys in there next to the page metadata.
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/hasdata-web-scraping)<a href="https://allmcps.com/mcp/hasdata-web-scraping"><img src="https://allmcps.com/api/badge/hasdata-web-scraping?style=directory" alt="HasData Web Scraping on AllMCPs" /></a>