ULTRA-ADVANCED PDF OCR Skill — Content stream scaling approach that produces the smallest possible searchable PDFs with perfect text selectability. Uses Tesseract's native PDF output and scales it to match original page dimensions via PDF content stream transforms.
Convert Markdown, HTML, plain text, images, and pandoc-supported formats to PDF with CJK font rendering, half-width numbers, and emoji support. Use this skill whenever the user asks to convert a file to PDF, generate a PDF from Markdown or HTML, export a PDF, or run "Convert to PDF" / "Generate PDF" / "Export PDF".
Download, split, and deeply read an academic PDF that is not available through Paperpile. Use when a long external PDF needs page-wise ingestion. For Paperpile items, use the Paperpile text-extraction route instead.
this skill when you need to read, inspect, or extract content from PDF files — especially when file content is NOT in your context and you need to read it from
Extract text, metadata, per-page density signals, layout, image boxes, OCR, and page PNGs from a PDF via the pdfvision CLI. Use when the input is a `.pdf` URL, a local PDF path, or a PDF another agent skill produced. Triggers on: 'read this pdf', 'extract from <file>.pdf', '.pdf', 'scan / slide / paper / form contents'.
Internal implementation skill invoked by /add-native for native PDF control workflows. Handles HTTPS and file URI PDF viewing with @microsoft/power-apps-native-pdf-viewer 0.2.9+.
Read PDF files by converting them to markdown. Use when the agent needs to read, analyze, or extract content from a PDF file. Handles caching (converts once, reads from cache thereafter) and automatically manages the Python virtual environment and dependencies.
Convert a local HTML file (including pages with embedded iframes) into a single-page tall PDF that visually matches the browser view (WYSIWYG, not paginated A4). Use whenever the user wants to turn an .html file on disk into a PDF for sharing with a colleague — phrasings like "output this html as pdf", "make a pdf of this prototype", "html to pdf", "render flow.html to pdf", "save this webpage as pdf", "screenshot this page as pdf", or any local .html mentioned alongside "send", "share", "export", or "for review". DO NOT use for fetching online URLs (different tool) or for paginated A4 print output (the user wants WYSIWYG).
使用 MinerU 精准 API (VLM) 将 PDF 转换为 Markdown。当用户需要解析 PDF 文档、将 PDF 转为可编辑的 Markdown 文本、提取 PDF 中的文字内容时使用。触发词包括"PDF 转 Markdown"、"解析 PDF"、"提取 PDF 文字"、"mineru"、"PDF to md"。
Use this skill to convert PDF pages into PNG or JPEG image files. Trigger when the user asks to render PDFs, export PDF pages as images, convert PDF to PNG, JPEG, or JPG, create page previews, or batch-convert PDF files. Do not use it for OCR, text extraction, PDF editing, or image-to-PDF conversion unless explicitly requested.
Generate a professional PDF report from a GEO audit using pandoc + Chrome headless. Converts GEO-AUDIT-REPORT.md into a styled, client-ready PDF with a cover page, color-coded score tables, severity-tagged findings, and a 90-day roadmap.
Render a Markdown or HTML document into a clean, branded, multi-page PDF via headless Chrome — the polished deliverable you'd hand a client, not a screenshot. Brand wrap (print-CSS + footer) is on by default; --no-brand for raw; --template for a custom shell. Never overwrites without --force. Use when the user says "make a PDF", "turn this into a PDF", "PDF this", "send them a PDF", "build a client doc", "export to PDF", "branded PDF".
Convert Markdown files (including Mermaid diagrams) to PDF using Chrome headless. Use when asked to "export to PDF", "convert md to pdf", "generate PDF from markdown", or when docs need to be shared as PDF. Supports Mermaid diagrams, tables, code blocks with syntax highlighting, and all standard markdown.
Plain text files in a repository that tell a coding agent how the project works: commands to run, conventions to follow and things to avoid. CLAUDE.md, AGENTS.md, cursor rules and skills are the common kinds.
CLAUDE.md or AGENTS.md?
CLAUDE.md is read by Claude Code. AGENTS.md is an open format that Codex, Cursor and other agents read. Many projects keep one and point the other at it.
What is a skill?
A folder with a SKILL.md that describes one capability, such as filling PDFs or reviewing code. The agent loads it only when the task calls for it.
Can I search my own team's files too?
Your agents already can, over MCP, limited to the files you're allowed to read. Searching them from this page is coming.