io.github.brenton-keller/pdf-text-table-ocr-extractorfiles

PDF Text, Table and OCR Extractor

Synced · awaiting check

PDF URLs to per-page text, tables as rows, Markdown, metadata and OCR for scanned pages.

Install

claude mcp add --transport http pdf-text-table-ocr-extractor https://mcp.apify.com/?tools=brenton8907/pdf-text-table-ocr-extractor

Our take

Extracts text, tables, Markdown, metadata, and OCR output from PDF URLs, covering both digital and scanned documents. It fits research, ingestion, and document-processing agents that need page-level structured content. The scope is practical, but remote-only operation and no visible package metadata mean reliability and data-handling expectations should be validated before production use.

reviewed by hand · 2026-09-09

Something wrong or dead here? Report it

Verification record

last verified
today
in registry since
2026-09-08
mcp-serverfiles

mcp-server

Verified · 3 days ago

Puter MCP enables AI tools to interact with Puter: manage files, websites, workers, and more

hand-reviewed43k starschecked 3 days ago

mcp-server-filesystemfiles

mcp-server-filesystem

Verified · 25 days ago

MCP server for filesystem access

hand-reviewed39k stars426 dl/wkchecked 25 days ago

desktop-commanderfiles

Desktop Commander

Verified · 2 days ago

MCP server for terminal commands, file operations, and process management

hand-reviewed9.5k stars39k dl/wkchecked 2 days ago

basic-memoryfiles

basic-memory

Verified · 7 days ago

Local-first knowledge management with bi-directional LLM sync via Markdown files.

hand-reviewed3.8k starschecked 7 days ago

IWE

Verified · 24 days ago

Markdown knowledge base as agent memory. Runs against the notes directory it is started in.

hand-reviewed1.5k stars237 dl/wkchecked 24 days ago

pdf-reader-mcpfiles

PDF Reader MCP

Verified · 21 days ago

Evidence-first PDF MCP. Agent Document Twin with citeable page+bbox evidence.

hand-reviewed894 stars2.2k dl/wkchecked 21 days ago