MinerU runs on your own machine — the client starts it, so there is no endpoint to ping. 618 installs a week from npm. Last commit 16 Sep 2026.
MinerU document parsing API — PDFs, images, DOCX, PPTX with OCR and batch processing.
We read the source, 22 h ago · rules 3dff92dd89df
A value the model can set ends up inside a file or shell call. That is not a flaw by itself — for a terminal server it is the job — but it is where things go wrong when it is not.
const out = execFileSync("mdls", ["-raw", "-name", "kMDItemNumberOfPages", params.file], { timeout: 10_000 }).toString().trim();
What this server is able to do. For an MCP server this is often the job itself — a terminal server runs commands because that is what it is for. Listed so you know what you are plugging in, not as an accusation.
import { readFileSync, createWriteStream, readdirSync, statSync, mkdirSync, existsSync, unlinkSync, rmSync, realpathSync, copyFileSync } from "node:fs";
Is this your server and something here is wrong? Tell us — corrections are free and do not require a plan.
That is not a flaw by itself — but it is where things go wrong when it is not the job. We re-read this code on every release. Watch it and you hear from us the day another one appears.
This server runs on your own machine — install it with the package manager and the client starts it for you. Package name taken from the official registry entry.
claude mcp add mineru -- npx -y mineru-mcp
{
"mcpServers": {
"mineru": {
"args": [
"-y",
"mineru-mcp"
],
"command": "npx"
}
}
}
[mcp_servers.mineru]
command = "npx"
args = ["-y", "mineru-mcp"]
{
"mcpServers": {
"mineru": {
"args": [
"-y",
"mineru-mcp"
],
"command": "npx"
}
}
}
{
"mcpServers": {
"mineru": {
"args": [
"-y",
"mineru-mcp"
],
"command": "npx"
}
}
}
Parse PDFs, images, doc, docx, ppt, pptx, xls, xlsx, html into Markdown using MinerU API.
Extract structured data from PDFs and documents with Suparse AI OCR.
Process video, audio, images, and documents with 86+ cloud media processing robots.
OCR for images and Korean ID documents
Extract text from documents, manipulate PDFs, and perform OCR on images.
Extract data, edit, convert, and parse PDF documents with OCR and AI
Convert documents, images and web pages to clean Markdown, for agents: single, URL and batch tools.
Compress OCR-heavy PDFs into dense packed images so agents can work with long visual documents.
Answers built from our own checks of this server.