The full upstream README, mirrored here for reference. Install config, tool schemas, adoption signals, and an original overview live on the Olostep MCP Server listing page.
Docker Hub npm version License: ISC
A Model Context Protocol (MCP) server implementation that integrates with Olostep for web scraping, content extraction, and search capabilities. To set up Olostep MCP Server, you need to have an API key. You can get the API key by signing up on the Olostep website.
There are multiple ways to connect to the Olostep MCP Server. Choose the one that best fits your workflow.
The simplest way — no local installation required. Connect directly to our hosted MCP server:
Authentication is done via a Bearer token in the Authorization header using your Olostep API key. See the Client Setup section below for configuration examples.
Pull and run the official Docker image:
If you prefer to build the image yourself from source:
Run without any installation using npx:
On Windows (PowerShell):
On Windows (CMD):
Or install globally:
The easiest way is to use the remote endpoint. Create or edit .cursor/mcp.json in your project root:
Alternative (local): Go to Cursor Settings > Features > MCP Servers, click "+ Add New MCP Server":
olostepcommandenv OLOSTEP_API_KEY=your-api-key npx -y olostep-mcpAdd this to your claude_desktop_config.json:
Alternative (Docker):
Or install via the Smithery CLI in your device terminal:
Add the remote endpoint to your Claude Code MCP configuration:
Alternative (local):
Add this to your ./codeium/windsurf/model_config.json:
Alternative (local):
Add this to your .vscode/mcp.json:
Alternative (local):
Option 1: One-Click Installation (Recommended)
Option 2: Manual Configuration
Add this to your Metorial MCP server configuration:
The Olostep tools will then be available in your Metorial AI chats.
OLOSTEP_API_KEY: Your Olostep API key (required)ORBIT_KEY: An optional key for using Orbit to route requests.scrape_website)Extract content from a single URL. Supports multiple formats and JavaScript rendering.
url_to_scrape: The URL of the website you want to scrape (required)output_format: Choose format (html, markdown, json, or text) - default: markdowncountry: Optional country code (e.g., US, GB, CA) for location-specific scrapingwait_before_scraping: Wait time in milliseconds before scraping (0-10000)parser: Optional parser ID for specialized extractionsearch_web)Search the Web for a given query and get structured results (non-AI, parser-based).
query: Search query (required)country: Optional country code for localized results (default: US)answers)Search the web and return AI-powered answers in the JSON structure you want, with sources and citations.
task: Question or task to answer using web data (required)json: Optional JSON schema/object or a short description of the desired output shapeanswer_id, object, task, result (JSON if provided), sources, createdbatch_scrape_urls)Scrape up to 10k URLs at the same time. Perfect for large-scale data extraction.
batch_id, status, total_urls, created_at, formats, country, parser, urlscreate_crawl)Start an async crawl that autonomously discovers and scrapes entire websites by following links. Returns a crawl_id — the crawl runs in the background and does not return content in this response. You must then call get_crawl_results with the crawl_id to poll status and retrieve the scraped pages (same two-step pattern as batch_scrape_urls + get_batch_results).
crawl_id, object, status, start_url, max_pages, created, formats, country, parserPair this call with
get_crawl_results— do not pass acrawl_idtoget_batch_results(crawls and batches are separate resources).
create_map)Get all URLs on a website. Extract all URLs for discovery and analysis.
map_id, object, url, total_urls, urls, search_query, top_nget_webpage_content)Retrieves webpage content in clean markdown format with support for JavaScript rendering.
url_to_scrape: The URL of the webpage to scrape (required)wait_before_scraping: Time to wait in milliseconds before starting the scrape (default: 0)country: Residential country to load the request from (e.g., US, CA, GB) (optional)get_website_urls)Search and retrieve relevant URLs from a website, sorted by relevance to your query.
url: The URL of the website to map (required)search_query: The search query to sort URLs by (required)get_batch_results)Retrieve the results of a previously submitted batch scrape job using its batch_id.
batch_id: The batch ID returned from batch_scrape_urls (required)batch_id, status (processing or completed), total_urls, completed_urls, items (array of scraped results per URL with url, custom_id, markdown_content, html_content, json_content, text_content, status, page_metadata)get_crawl_results)Retrieve the status and scraped pages for an async crawl started with create_crawl. This is the required companion to create_crawl — create_crawl only kicks off the job and returns a crawl_id; this tool is how you actually fetch the discovered pages and their content.
crawl_id: The crawl ID returned from create_crawl (required)formats: Array of formats to retrieve per page — markdown, html, json, text (default: ["markdown"])items_limit: Max pages to retrieve content for, 1–100 (default: 20)cursor: Pagination cursor into the list of discovered pages (default: 0)search_query: Optional filter to rank/select pages by relevance to a querycrawl_id, status (in_progress), pages_completed, pages_total, and a message prompting you to call again in ~10 seconds.crawl_id, status (completed), pages_returned, next_cursor, has_more, and a pages array where each entry has url, custom_id, and the requested content fields (markdown_content, html_content, json_content, text_content).The server provides robust error handling:
Example error response:
The MCP server is available as a Docker image:
[olostep/mcp-server](https://hub.docker.com/r/olostep/mcp-server)mcp/olostep (coming soon - enhanced security with signatures & SBOMs)ghcr.io/olostep/olostep-mcp-serverThe Olostep MCP Server is being added to Docker Desktop's official MCP Toolkit, which means users will be able to:
Status: Submission in progress to Docker MCP Registry
linux/amd64linux/arm64ISC License