Simplescraper extracts structured data from websites using AI, discovers page URLs from sitemaps, converts web pages to Markdown, captures screenshots, and automates reusable scraping recipes for bulk data collection and monitoring.
Encrypted at rest, isolated from the model
Resolved from an AES-256-GCM vault at the moment of the call and attached to the request — the model never sees the secrets.
Try asking
Get the most recent scraped results for a recipe. Returns the existing data from the last completed run without scraping again.
List previous scrape runs for a recipe. Each entry has results_id, scrape timestamp, and page count. Returns run metadata only, not the scraped data itself.
List the user's scrape recipes, with optional filters by domain, name keyword, or recent activity, and sorting by creation date, last-run, or name.
Run a scrape recipe to extract fresh data from its configured URL. Supports async mode, source URL override, markdown extraction, and batch (crawler) mode.
Create a new scrape recipe with a name, URL, and selectors.
Retrieve full details of a specific scrape recipe: name, URL, creation date, last run time, and selectors.
Update an existing scrape recipe's name, URL, or selectors. Supports 'replace' (default) or 'merge' modes for selector updates.
Fetch the scraped data from a specific run by results_id. Supports ordering (by index or timestamp) and pagination via limit and offset.
Extract specific named fields from one page using AI. Pass a comma-separated schema like 'product name, price, sku' and get those fields back as a structured object, plus the CSS selectors that matched them.
Discover all page URLs on a website by parsing its sitemap (and the sitemap index if present). Returns a deduplicated list of page URLs.
Fetch one page and return its body as clean Markdown. Renders JavaScript before extraction. Returns markdown plus metadata (title, word count).
Add or replace the list of batch (crawler) URLs for a recipe. 'append' adds to the existing list, 'replace' overwrites it.
Capture a screenshot of a web page (1 credit per request). Returns JSON with a hosted screenshot URL by default. Supports full-page capture, custom viewport, format/quality, delays, and popup/ad hiding. Defaults output to 'url' so the response stays small; 'binary' returns raw image bytes and is not suited to an MCP text response.
One endpoint, the same key, whichever client you use.
~/Library/Application Support/Claude/claude_desktop_config.json (Mac) · %APPDATA%\Claude\claude_desktop_config.json (Windows)
Replace API_KEY with your own key.
Already have an "mcpServers" section in your config? Just add the server entry inside it.
Discovery, routing, credentials, tool scoping and execution logs all happen at the gateway→connections stay ACTIVE with no work from you
Simplescraper MCP runs through a gateway that holds the credentials, scopes the access and records every call.
Managed auth, hosted MCP servers, and every Gmail tool your agent needs.
Free to start.