doc.page PDF Extraction is answering right now. Last checked 9 min ago. It exposes 7 tools.
Extract PDFs to Markdown, RAG chunks and cited tables; publish tracked Doc Links with read stats.
Today is the operative word: we check doc.page PDF Extraction every 15 minutes and re-read its code on every release. Watch it and you find out the day that stops being true.
Endpoint below is the one we actually reach during checks — not the one copied from a README. Last verified 9 min ago.
claude mcp add pdf-extract --transport http https://doc.page/api/mcp
{
"mcpServers": {
"pdf-extract": {
"url": "https://doc.page/api/mcp"
}
}
}
[mcp_servers.pdf-extract]
url = "https://doc.page/api/mcp"
{
"mcpServers": {
"pdf-extract": {
"url": "https://doc.page/api/mcp"
}
}
}
{
"mcpServers": {
"pdf-extract": {
"url": "https://doc.page/api/mcp"
}
}
}
Read directly from the server with tools/list, grouped by what they act on.
If a tool disappears, we record the date.
create_doc_link
get_doc_link_stats
list_doc_links
get_chunks
extract_pdf
revoke_doc_link
list_tables
| URL | Transport | State | Latency | Checked |
|---|---|---|---|---|
| https://doc.page/api/mcp | streamable-http | answering | 295 ms | 9 min ago |
Universal web scraper with LLM-ready markdown, RAG chunking, PDF/DOCX support.
Text extraction, keyword extraction, language detection, and chunking for RAG
Sync public web sources into cited context packs for AI agents, RAG, and MCP clients.
Read-only local-first MCP server for private Markdown, PDF, and Tika-backed search on Windows.
Read-only local-first MCP server for private Markdown, PDF, and Tika-backed search on Windows.
Read markdown with inlined images and index headings for agentic RAG workflows.
Reasoning-based RAG system for chatting with long PDFs. Supports local and online files.
Inject, parse, and strip [N] citation markers in RAG outputs.
Answers built from our own checks of this server.