
PDF Text Extractor (Apify Actor)
io.github.truhojunbot-techv1.0.0Updated Oct 8, 2026
Extract text, tables and RAG chunks from public PDF URLs; pay-per-event via your Apify account.
Overview
AI-generated overview
Extracts text, tables and RAG-ready chunks from public PDF URLs through an Apify-hosted remote MCP endpoint.
- What it does
- This remote MCP server exposes an Apify Actor that pulls text, tables and RAG chunks out of public PDF files reachable by URL. The assistant sends a PDF link and receives the extracted content back for further use. It runs on Apify's hosted MCP endpoint rather than on the user's machine.
- When to use it
- Useful when an assistant needs to read or reuse the contents of public PDFs, for example to feed document text into a retrieval or RAG pipeline, or to pull tables out of reports and papers. It is not intended for private or non-public documents.
- Requirements
- A remote MCP client pointed at the Apify MCP endpoint. Authentication is optional: an Authorization header with a Bearer APIFY_TOKEN, or OAuth sign-in through the client if no token is supplied. Runs are billed to the caller's Apify account, so an Apify account is needed.
Before you install
Runs are billed per event to the caller's Apify account, so usage costs money. The Authorization header carries an Apify token secret and should be handled carefully. PDF URLs and extracted content are processed by Apify's hosted service, a third party.
Installation
In SourceWeft
- Open PDF Text Extractor (Apify Actor) in the dashboard and add it to a workspace.
- Enable the server for the chats that should use its tools.
Web executable via Streamable HTTP. Remote servers run from the web runtime once configured in a workspace.
Other MCP clients
Add this to your client's mcpServers config.
{
"mcpServers": {
"pdf-text-extractor": {
"type": "http",
"url": "https://mcp.apify.com/?tools=gochujang/pdf-text-extractor"
}
}
}Tools
0Tool metadata has not been indexed yet.
Version history
1- v1.0.0LatestOct 8, 2026

