ScrapingBot
io.scrapingbotv1.0.0更新於 Oct 4, 2026
Public web data for AI agents: scrape any page, plus Google, TikTok, Instagram and Amazon as JSON.
概覽
讓助理抓取任意網頁,並以 JSON 取得 Google、TikTok、Instagram 與 Amazon 的公開資料,透過託管端點接入。
- 功能
- 這是一個託管型 MCP 伺服器,透過 Streamable HTTP 提供 23 個工具。它能把任意網頁抓取為 markdown 或 HTML,可選 JavaScript 渲染、截圖與瀏覽器操作,並可用 AI 擷取結構化欄位。其他工具涵蓋 Google 搜尋、圖片、影片、新聞、購物、地點、地圖與評論;Instagram 個人檔案、媒體、粉絲與搜尋;TikTok 影片、使用者、留言與搜尋;以及 Amazon 搜尋、商品詳情、排行榜與搜尋建議。另有工具用於列出能力與輪詢抓取工作。
- 適用情境
- 當助理需要即時公開網頁資料時使用:抓取並解析頁面、執行瀏覽器操作、用 AI 擷取欄位,或查詢公開的 Google、TikTok、Instagram 與 Amazon 資料,而不必自建爬蟲或代理。它是按額度計費的付費服務,適合偶爾或中等規模的資料蒐集,不適合大規模持續爬取。若只處理本機檔案或私有資料則無需使用。
- 執行需求
- 遠端託管的 MCP 端點 Streamable HTTP,無需安裝。需要 ScrapingBot 帳號與 API 金鑰,透過 x-api-key 標頭(或 Authorization: Bearer,或僅接受 URL 的用戶端使用 api_key 查詢參數)傳送。需要能連線 scrapingbot.io。註冊免費贈送 100 額度,不需信用卡。
安裝
在 SourceWeft 中
- 開啟 儀表板中的 ScrapingBot,將其新增到工作區。
- 為需要使用其工具的對話啟用該服務。
Web executable,透過 Streamable HTTP。 遠端服務在工作區中設定後即可從網頁執行環境執行。
其他 MCP 客戶端
把它新增到你客戶端的 mcpServers 設定中。
{
"mcpServers": {
"scrapingbot": {
"type": "http",
"url": "https://scrapingbot.io/api/mcp"
}
}
}README
ScrapingBot
One API key for public web data: scrape any page (plain or JavaScript-rendered), pull fields out with AI, and get public TikTok, Instagram, Google and Amazon data as JSON. A hosted MCP server gives AI agents the same data as tools.
100 free credits when you sign up. No credit card. Get your API key
Docs · Pricing · MCP server · Examples · Agent Skills
Quickstart
That call cost 1 credit. Failed requests are refunded automatically.
Base URL: https://scrapingbot.io/api/v1
Auth: x-api-key: YOUR_API_KEY header (or Authorization: Bearer YOUR_API_KEY). Keep the key on the server.
Endpoints
The data APIs (TikTok, Instagram, Google, Amazon) are all POST with a JSON body naming the endpoint and its params:
Full parameters and response shapes: docs.
How charging works
- Data APIs charge only for a successful
200. Any error is refunded. - Website API charges when the page answers (2xx/3xx, or a real 400/404 from the site, reported as
fault: "user"). Timeouts, other 4xx such as 403 and 429, 5xx and errors on our side are refunded (credits_used: 0). - Requests rejected before they run (missing parameter, bad key, concurrency limit) cost nothing.
Limits
Limits are on requests in flight at once, not per minute. The free plan allows 1 concurrent request; paid plans allow 10 to 200. Going over returns 429 right away (free); retry with a short, growing delay. Timeouts: 45 s for scraping and Instagram, 30 s for TikTok and Amazon, 15 s for Google.
MCP server for AI agents
Hosted at https://scrapingbot.io/api/mcp (Streamable HTTP, stateless). Nothing to install. Authenticate with your API key in the x-api-key header, or Authorization: Bearer.
Claude Code
Add --scope user to use it in every project. Type /mcp in a session to see the tools.
Claude Desktop (and Claude on the web)
Open Settings → Connectors → Add custom connector, name it ScrapingBot, and paste:
Then turn it on from the tools menu in a chat. Connectors take a URL only, so the key goes in the URL: keep that URL private, like a password.
Cursor
~/.cursor/mcp.json (or .cursor/mcp.json in a project):
VS Code
.vscode/mcp.json in your project, then pick the ScrapingBot tools in Copilot's agent mode:
Any other MCP client
URL https://scrapingbot.io/api/mcp, header x-api-key: YOUR_API_KEY (or Authorization: Bearer YOUR_API_KEY). If the client only takes a URL, append ?api_key=YOUR_API_KEY. When calling it by hand, send Accept: application/json, text/event-stream.
Tools
Tool calls are ordinary API requests: same credits, same concurrency slots, same refunds.
Agent Skills
skills/ holds Agent Skills that teach an agent to call the REST API directly with SCRAPINGBOT_API_KEY: which endpoint for which task, the parameters, the response fields worth reading, error handling and costs.
Claude Code: copy a folder into ~/.claude/skills/ (or .claude/skills/ in a project). Set SCRAPINGBOT_API_KEY in the environment the agent runs in.
Examples
Small, runnable scripts in examples/, each in curl, Python (requests) and Node (built-in fetch, Node 18+):
Python needs pip install requests; the curl scripts that build JSON use jq. Each language folder has a tiny shared client (scrapingbot.py, scrapingbot.mjs) that retries refunded failures (408, 429, 5xx) with backoff.
Pricing
Free: 100 credits on sign-up, 1 concurrent request. Paid plans start at $49.99/month for 275,000 credits and 10 concurrent requests. See pricing.
Responsible use
ScrapingBot returns publicly available data. Respect each site's terms and applicable law, and handle any personal data you collect accordingly.
License
The examples and skills in this repository are MIT licensed. See LICENSE. Use of the ScrapingBot API is subject to the terms at scrapingbot.io.
來源:README.md,提交 349f2bf
工具
0版本歷史
1- v1.0.0最新Oct 4, 2026
