firecrawl/firecrawl-knowledge-ingest
Ingest public or authenticated knowledge bases and docs portals with Firecrawl browser. Use for JS-heavy docs, login-gated portals, paginated help centers, support knowledge bases, or structured JSON/markdown extraction from documentation sites.
npx skills add https://github.com/firecrawl/firecrawl-workflows --skill firecrawl-knowledge-ingest
Use this when a docs portal needs browser navigation, auth, pagination, or JS rendering.
Infer the portal URL, output format, auth needs, and page limit from context. If the portal is clear, proceed immediately.
Ask at most 1-3 concise questions only if blocked, such as the portal URL, whether authentication is required, or the desired output format.
Use Firecrawl browser to:
Try Firecrawl map as a supplement for public URLs, but use browser navigation for auth-gated or JS-heavy content.
# Knowledge Ingest: [Portal]
## Summary
[Pages extracted, sections covered, limitations]
## Output
[JSON/markdown/merged file path or content]
## Sections
[Section names and article counts]
## Failed Or Restricted Pages
[Any access/loading issues]
## Sources
[URLs extracted]
## Rerun Inputs
workflow: firecrawl-knowledge-ingest
url: [portal url]
format: [json/markdown/merged]
max_pages: [number]
Use source, url, extractedAt, totalArticles, and sections[] with article title, url, section, content, and metadata.
Take firecrawl/firecrawl-knowledge-ingest from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.