pdfmux runs on your own machine — the client starts it, so there is no endpoint to ping. 598 installs a week from pypi. Last commit 12 Sep 2026.
PDF-to-Markdown extraction that audits its own output and flags any extractor's silent drops.
We read the source, 19 h ago · rules 3dff92dd89df
What this server is able to do. For an MCP server this is often the job itself — a terminal server runs commands because that is what it is for. Listed so you know what you are plugging in, not as an accusation.
этот файл ставится пользователю, но в репозитории его нет
const version = execSync(`${cmd} --version 2>&1`, { encoding: "utf8" });
Is this your server and something here is wrong? Tell us — corrections are free and do not require a plan.
We found places where it runs commands, builds paths or queries from values it is given. None of that is a flaw by itself — it becomes one when the code changes, and code changes quietly between releases. We re-read it on every one.
This server runs on your own machine — install it with the package manager and the client starts it for you. Package name taken from the official registry entry.
claude mcp add pdfmux -- uvx pdfmux
{
"mcpServers": {
"pdfmux": {
"args": [
"pdfmux"
],
"command": "uvx"
}
}
}
[mcp_servers.pdfmux]
command = "uvx"
args = ["pdfmux"]
{
"mcpServers": {
"pdfmux": {
"args": [
"pdfmux"
],
"command": "uvx"
}
}
}
{
"mcpServers": {
"pdfmux": {
"args": [
"pdfmux"
],
"command": "uvx"
}
}
}
This one needs environment variables set before it will start:
GEMINI_API_KEY (Optional Google AI API key. Enables the paid Gemini Flash fallback for pages where all local backends report low confidence. pdfmux runs 100% locally without this.).
The author declared them in the registry entry; get the values from the project itself.
Screenshot, PDF, and markdown extraction from any URL via Claude Desktop and Claude Code.
High-fidelity PDF to structured Markdown conversion and document field extraction.
Render HTML, Markdown, or URLs to images, PDF, or branded artifacts; extract and watch pages.
Universal web content extraction — any URL to LLM-ready markdown. HTML, YouTube, PDF, DOCX.
PDF URLs to per-page text, tables as rows, Markdown, metadata and OCR for scanned pages.
Extract text, tables and metadata from every PDF linked in a dataset, CSV or Google Sheet.
Extract PDFs to Markdown, RAG chunks and cited tables; publish tracked Doc Links with read stats.
Convert PDF, DOCX, HTML, Markdown, and Text for AI assistant context injection.
Answers built from our own checks of this server.