ScrapingBot

io.scrapingbotv1.0.0Updated Oct 4, 2026

Public web data for AI agents: scrape any page, plus Google, TikTok, Instagram and Amazon as JSON.

VerifiedStreamable HTTPWeb executableWeb Search & ScrapingData & AnalyticsBrowser Automation

Overview

AI-generated overview

Lets an assistant scrape any web page and pull public Google, TikTok, Instagram and Amazon data as JSON through a hosted endpoint.

What it does
A hosted MCP server exposing 23 tools over Streamable HTTP. It scrapes any page as markdown or HTML with optional JavaScript rendering, screenshots and browser scenarios, and can extract structured fields with AI. Separate tools cover Google search, images, videos, news, shopping, places, maps and reviews; Instagram profiles, media, followers and search; TikTok videos, users, comments and search; and Amazon search, product details, rankings and suggestions. Utility tools list capabilities and poll scrape jobs.
When to use it
Use it when an assistant needs live public web data: fetching and parsing pages, running browser actions, extracting fields with AI, or querying public Google, TikTok, Instagram and Amazon data without building your own scrapers or proxies. It is a paid, credit-metered service, so it suits occasional or moderate data collection rather than high-volume crawling. Not needed if you only work with local files or private data.
Requirements
Remote hosted MCP endpoint at over Streamable HTTP; nothing to install. Needs a ScrapingBot account and API key, sent in the x-api-key header (or Authorization: Bearer, or an api_key query parameter for URL-only clients). Network access to scrapingbot.io. Free sign-up gives 100 credits with no credit card.
Before you install
Requires a ScrapingBot API key (x-api-key header, or Authorization Bearer; some clients put it in the URL as api_key, which the README says to keep private like a password). Calls consume paid credits: 1 per plain scrape, 5 rendered, 10 premium proxy, 75 stealth proxy, +5 for AI extraction, 5 per Instagram call, 10 per Google or Amazon call; free plan has 100 credits and 1 concurrent request, paid plans start at $49.99/month. Tool calls are ordinary API requests and are charged the same. Scraped

Installation

In SourceWeft

  1. Open ScrapingBot in the dashboard and add it to a workspace.
  2. Enable the server for the chats that should use its tools.

Web executable via Streamable HTTP. Remote servers run from the web runtime once configured in a workspace.

Other MCP clients

Add this to your client's mcpServers config.

{
  "mcpServers": {
    "scrapingbot": {
      "type": "http",
      "url": "https://scrapingbot.io/api/mcp"
    }
  }
}

README

ScrapingBot

One API key for public web data: scrape any page (plain or JavaScript-rendered), pull fields out with AI, and get public TikTok, Instagram, Google and Amazon data as JSON. A hosted MCP server gives AI agents the same data as tools.

100 free credits when you sign up. No credit card. Get your API key

Docs · Pricing · MCP server · Examples · Agent Skills


Quickstart

bash
export SCRAPINGBOT_API_KEY="YOUR_API_KEY"
curl "https://scrapingbot.io/api/v1/scrape?url=https://example.com" \  -H "x-api-key: $SCRAPINGBOT_API_KEY"
json
{  "success": true,  "url": "https://example.com",  "html": "<!doctype html><html lang=en><head>…",  "status": 200,  "duration": "0.79",  "credits_used": 1,  "job_id": "…"}

That call cost 1 credit. Failed requests are refunded automatically.

Base URL: https://scrapingbot.io/api/v1 Auth: x-api-key: YOUR_API_KEY header (or Authorization: Bearer YOUR_API_KEY). Keep the key on the server.

Endpoints

The data APIs (TikTok, Instagram, Google, Amazon) are all POST with a JSON body naming the endpoint and its params:

bash
curl -X POST "https://scrapingbot.io/api/v1/tiktok" \  -H "x-api-key: $SCRAPINGBOT_API_KEY" \  -H "Content-Type: application/json" \  -d '{"endpoint": "/user/info", "params": {"unique_id": "tiktok"}}'
APIRouteendpoint valuesCredits
WebsiteGET or POST /api/v1/scrapeone route; options: render_js, screenshot, js_scenario, wait_for, premium_proxy, stealth_proxy, …1 plain · 5 rendered · 10 premium proxy · 75 stealth proxy
AI extraction/api/v1/scrape + ai_query or ai_schemareturns ai_result JSON+5 on top of the page
Job lookupGET /api/v1/job/:job_idre-fetch a scrape result (kept 24 hours)free
TikTokPOST /api/v1/tiktok/ (video), /user/info, /user/posts, /user/followers, /user/following, /user/search, /feed/search, /music/info, /music/posts, /comment/list, /comment/reply1
InstagramPOST /api/v1/instagram/user/by_username, /user/by_id, /medias/by_user_id, /reels/by_user_id, /medias/tagged_by_user_id, /stories/by_username, /followers/by_user_id, /following/by_user_id, /media/by_shortcode, /media/by_url, /comments/media_comments_by_id, /comments/replies, /search/users_by_keyword, /search/hashtags_by_keyword, /search/places_by_keyword, /search/global, /search/posts, /media/shortcode_to_id, /media/id_to_shortcode5
GooglePOST /api/v1/google/search, /images, /videos, /news, /shopping, /places, /maps, /reviews10
AmazonPOST /api/v1/amazon/search, /product-details, /products (up to 20 ASINs), /autocomplete, /best-sellers, /new-releases, /product-category-list, /deals-v2, /seller-profile, /seller-products10 (/products: 10 per product returned)
ChatGPTPOST /api/v1/chatgptbody {"prompt": "…"}10
MCPPOST /api/mcp23 tools over Streamable HTTPsame as the API behind each tool

Full parameters and response shapes: docs.

How charging works

  • Data APIs charge only for a successful 200. Any error is refunded.
  • Website API charges when the page answers (2xx/3xx, or a real 400/404 from the site, reported as fault: "user"). Timeouts, other 4xx such as 403 and 429, 5xx and errors on our side are refunded (credits_used: 0).
  • Requests rejected before they run (missing parameter, bad key, concurrency limit) cost nothing.

Limits

Limits are on requests in flight at once, not per minute. The free plan allows 1 concurrent request; paid plans allow 10 to 200. Going over returns 429 right away (free); retry with a short, growing delay. Timeouts: 45 s for scraping and Instagram, 30 s for TikTok and Amazon, 15 s for Google.

StatusMeaningCharged
400Missing or invalid parameter, unsupported endpointNo (except a target page's own 400 on the Website API)
401Missing or invalid API keyNo
402Not enough credits; the message says how many are neededNo
404Website API: the page doesn't exist. Data APIs: profile, post or product not foundWebsite API only
408Timed outNo
429Concurrency limit reachedNo
5xxThe site or a data source failedNo

MCP server for AI agents

Hosted at https://scrapingbot.io/api/mcp (Streamable HTTP, stateless). Nothing to install. Authenticate with your API key in the x-api-key header, or Authorization: Bearer.

Claude Code

bash
claude mcp add --transport http scrapingbot https://scrapingbot.io/api/mcp \  --header "x-api-key: YOUR_API_KEY"

Add --scope user to use it in every project. Type /mcp in a session to see the tools.

Claude Desktop (and Claude on the web)

Open Settings → Connectors → Add custom connector, name it ScrapingBot, and paste:

https://scrapingbot.io/api/mcp?api_key=YOUR_API_KEY

Then turn it on from the tools menu in a chat. Connectors take a URL only, so the key goes in the URL: keep that URL private, like a password.

Cursor

~/.cursor/mcp.json (or .cursor/mcp.json in a project):

json
{  "mcpServers": {    "scrapingbot": {      "url": "https://scrapingbot.io/api/mcp",      "headers": { "x-api-key": "YOUR_API_KEY" }    }  }}

VS Code

.vscode/mcp.json in your project, then pick the ScrapingBot tools in Copilot's agent mode:

json
{  "servers": {    "scrapingbot": {      "type": "http",      "url": "https://scrapingbot.io/api/mcp",      "headers": { "x-api-key": "YOUR_API_KEY" }    }  }}

Any other MCP client

URL https://scrapingbot.io/api/mcp, header x-api-key: YOUR_API_KEY (or Authorization: Bearer YOUR_API_KEY). If the client only takes a URL, append ?api_key=YOUR_API_KEY. When calling it by hand, send Accept: application/json, text/event-stream.

Tools

AreaToolsCredits
Any websitescrapeWebsite (markdown or HTML, optional screenshot), extractStructuredData (AI fields as JSON), runBrowserScenario (click, fill, scroll, wait)1 / 5 rendered; +5 for AI
GooglegoogleSearch (web, images, videos, news, shopping, places, maps), googleReviews10
InstagraminstagramUser, instagramSearch, instagramMedia, instagramFollowers5
TikToktiktokVideo, tiktokUser, tiktokSearch, tiktokComments, tiktokFollowers1
AmazonamazonSearch, amazonProduct, amazonProducts, amazonSuggestions, amazonRankings10
UtilitylistCapabilities, getScrapeJob, pollJobUntilDone (free), providerRequest (any endpoint above)

Tool calls are ordinary API requests: same credits, same concurrency slots, same refunds.

Agent Skills

skills/ holds Agent Skills that teach an agent to call the REST API directly with SCRAPINGBOT_API_KEY: which endpoint for which task, the parameters, the response fields worth reading, error handling and costs.

SkillUse it for
web-scrapingFetching any page, rendering JavaScript, browser actions, screenshots, AI extraction
google-searchGoogle web, images, videos, news, shopping, places, maps and reviews
tiktok-dataPublic TikTok videos, profiles, posts, sounds, comments and search
instagram-dataPublic Instagram profiles, posts, reels, comments and search
amazon-productsAmazon search, product details, batches, rankings, suggestions and deals

Claude Code: copy a folder into ~/.claude/skills/ (or .claude/skills/ in a project). Set SCRAPINGBOT_API_KEY in the environment the agent runs in.

Examples

Small, runnable scripts in examples/, each in curl, Python (requests) and Node (built-in fetch, Node 18+):

Python needs pip install requests; the curl scripts that build JSON use jq. Each language folder has a tiny shared client (scrapingbot.py, scrapingbot.mjs) that retries refunded failures (408, 429, 5xx) with backoff.

bash
export SCRAPINGBOT_API_KEY="YOUR_API_KEY"sh examples/curl/scrape.sh https://example.compython examples/python/google_search.py "best espresso machine"node examples/node/tiktok_user.mjs tiktok

Pricing

Free: 100 credits on sign-up, 1 concurrent request. Paid plans start at $49.99/month for 275,000 credits and 10 concurrent requests. See pricing.

Responsible use

ScrapingBot returns publicly available data. Respect each site's terms and applicable law, and handle any personal data you collect accordingly.

License

The examples and skills in this repository are MIT licensed. See LICENSE. Use of the ScrapingBot API is subject to the terms at scrapingbot.io.

Source: README.md at commit 349f2bf

Tools

0
Tool metadata has not been indexed yet.

Version history

1
  1. v1.0.0LatestOct 4, 2026