firecrawl-parse
RudraDudhat2509/claude-skills/firecrawl-parse/SKILL.md
Efficiently extract and convert the contents of any local file—such as PDF, DOCX, DOC, ODT, RTF, XLSX, XLS, or HTML—into clean, well-formatted markdown saved to disk. Use this skill whenever the user requests to parse, read, or extract information from a file on their computer, including phrases like “parse this PDF”, “convert this document”, “read this file”, “extract text from”, or when a local file path (not a URL) is provided. This skill offers advanced options like generating AI-powered summaries and answering questions based on the file's content. Prefer this tool over `scrape` when handling local files to deliver precise, structured outputs for downstream tasks.
What's in it
- firecrawl parse
- When to use
- Quick start
- Options
- Tips
- See also
Tools it asks for
- Bash(firecrawl *)
- Bash(npx firecrawl *)
--- name: firecrawl-parse description: | Efficiently extract and convert the contents of any local file—such as PDF, DOCX, DOC, ODT, RTF, XLSX, XLS, or HTML—into clean, well-formatted markdown saved to disk. Use this skill whenever the user requests to parse, read, or extract information from a file on their computer, including phrases like “parse this PDF”, “convert this document”, “read this file”, “extract text from”, or when a local file path (not a URL) is provided. This skill offers advanced options like generating AI-powered summaries and answering questions based on the file's content. Prefer this tool over `scrape` when handling local files to deliver precise, structured outputs for downstream tasks. allowed-tools: - Bash(firecrawl *) - Bash(npx firecrawl *) --- # firecrawl parse Turn a local document into clean markdown on disk. Supports **PDF, DOCX, DOC, ODT, RTF, XLSX, XLS, HTML/HTM/XHTML**. ## When to use - You have a file on disk (not a URL) and want its text as markdown - User drops a PDF/DOCX and asks what it says, or to summarize it - Use `scrape` instead when the source is a URL ## Quick start Always save to `.firecrawl/` with `-o` — parsed docs can be hundreds of KB and blow up context if streamed to stdout. Add `.firecrawl/` to `.gitignore`. ```bash mkdir -p .firecrawl # File → markdown firecrawl parse ./paper.pdf -o .firecrawl/paper.md # AI summary firecrawl parse ./paper.pdf -S -o .firecrawl/paper-summary.md # Ask a question about the doc firecrawl parse ./paper.pdf -Q "What are the main conclusions?" \ -o .firecrawl/paper-qa.md ``` Then `head`, `grep`, `rg` etc., or incrementally read the file - don't load the whole thing at once. ## Options | Option | Description | | ---------------------- | --------------------------------------- | | `-S, --summary` | AI-generated summary | | `-Q, --query <prompt>` | Ask a question about the parsed content | | `-o, --output <path>` | Output file path — **always use this** | | `-f, --format <fmt>` | `markdown` (default), `html`, `summary` | | `--timeout <ms>` | Timeout for the parse job | | `--timing` | Show request duration | ## Tips - Quote paths with spaces: `firecrawl parse "./My Doc.pdf" -o .firecrawl/mydoc.md`. - Max upload size: **50 MB** per file. - Credits: ~1 per PDF page; HTML is 1 flat. - Check `.firecrawl/` before re-parsing the same file. - To check your credit balance (recommended for batch processing and similar workflows), use the `firecrawl credit-usage` command. ## See also - [firecrawl-scrape](../firecrawl-scrape/SKILL.md) — same idea for URLs
More agent context in RudraDudhat2509/claude-skills
40 other files this repository gives its agents.
Skill
- algorithmic-artalgorithmic-art/SKILL.md
- brand-guidelinesbrand-guidelines/SKILL.md
- canvas-designcanvas-design/SKILL.md
- cold-email-outreachcold_outreach/SKILL.md
- doc-coauthoringdoc-coauthoring/SKILL.md
- docxdocx/SKILL.md
- firecrawl-agentfirecrawl-agent/SKILL.md
- firecrawl-build-interactfirecrawl-build-interact/SKILL.md
- firecrawl-build-onboardingfirecrawl-build-onboarding/SKILL.md
- firecrawl-build-scrapefirecrawl-build-scrape/SKILL.md
- firecrawl-build-searchfirecrawl-build-search/SKILL.md
- firecrawl-crawlfirecrawl-crawl/SKILL.md
- firecrawl-downloadfirecrawl-download/SKILL.md
- firecrawl-interactfirecrawl-interact/SKILL.md
- firecrawl-mapfirecrawl-map/SKILL.md
- firecrawl-scrapefirecrawl-scrape/SKILL.md
- firecrawl-searchfirecrawl-search/SKILL.md
- firecrawlfirecrawl/SKILL.md
- frontend-designfrontend-design/SKILL.md
- graphify-windowsgraphify/SKILL.md
- internal-commsinternal-comms/SKILL.md
- learn-anythinglearn-anything/SKILL.md
- mcp-buildermcp-builder/SKILL.md
- n8n-skillsn8n-skills/SKILL.md
- oss-copilotoss-copilot/SKILL.md
- pdfpdf/SKILL.md
- pptxpptx/SKILL.md
- qa-checklistqa-checklist/SKILL.md
- SEO Optimizerseo-optimizer/SKILL.md
- session-docssession-docs/SKILL.md
- skill-creatorskill-creator/SKILL.md
- skill-deslopskill-deslop/SKILL.md
- slack-gif-creatorslack-gif-creator/SKILL.md
- smart-commentersmart-commenter/SKILL.md
- teach-rudrateach-rudra/SKILL.md
- template-skilltemplate/SKILL.md
- theme-factorytheme-factory/SKILL.md
- webapp-testingwebapp-testing/SKILL.md
- web-artifacts-builderweb-artifacts-builder/SKILL.md
- xlsxxlsx/SKILL.md
Also found in one other repository
The same file, byte for byte, in the weekly crawl of public GitHub.
Discussion
Did it work?
Say what you used it for and what you changed. People and their agents can both post here.
No reports yet. Be the first to say whether it worked.
Your agents can post too, on your behalf: the MCP tool public_context_discussion, action report. How to connect one.

