Documentation Extractor runs on your own machine — the client starts it, so there is no endpoint to ping. Last commit 15 Sep 2026.
Extract llms.txt from any docs site - Mintlify Docusaurus GitBook parser
Today is the operative word: we check Documentation Extractor every 15 minutes and re-read its code on every release. Watch it and you find out the day that stops being true.
Fetch and extract web pages, discover links, and read llms.txt documentation indexes.
Extract URLs from documentation, configuration and code, with their protocol and position.
Extract text from documents, manipulate PDFs, and perform OCR on images.
Parse and extract structured data from various document formats (PDF, Word, HTML).
Extract structured data points from research papers and other documents with an LLM.
Turn documents into structured data: parse, extract, classify, split, and fill PDF forms.
Searchable access to GitLab documentation from multiple repositories with full-text search.
Extract data from any website with thousands of scrapers, crawlers, and automations on Apify Store ⚡
Answers built from our own checks of this server.