Use when the user needs PDF generation, manipulation, form filling, table extraction, OCR, merging, splitting, watermarking, or metadata handling. Trigger conditions: generate PDF reports, extract text or tables from PDFs, fill PDF forms programmatically, merge or split PDF files, add watermarks, OCR scanned documents, read or write PDF metadata, convert HTML to PDF.
Converts a Markdown document to PDF with full Mermaid diagram rendering (charts render as visuals, not code). Use this skill whenever the user wants to export a use case, sequence diagram, or review document to PDF — or any .md file in the project. Triggered by phrases like "export to pdf", "convert to pdf", "generate pdf", "save as pdf", or "/md-to-pdf". Works with UC-, SEQ-, REVIEW- prefixed files and any other Markdown file.
Create PDF documents from markdown with proper Chinese font support using weasyprint. This skill should be used when converting markdown to PDF, generating formal documents (legal, trademark filings, reports), or when Chinese typography is required. Triggers include "convert to PDF", "generate PDF", "markdown to PDF", or any request for creating printable documents.
Convert PDF files to high-fidelity Markdown using Docling (IBM). Preserves tables (TableFormer ML), images (extracted as PNG or embedded base64), formulas, headings, and reading order. Use when the user wants to convert a PDF to .md with high accuracy, especially when tables or images are present. Triggers on "/pdf-to-md", "converti pdf in markdown", "pdf to md", "estrai testo da pdf", "trasforma pdf in md".
Convert one or more Markdown files to clean, print-styled PDFs. Use when the user asks for "a PDF version of <doc>.md", "export this markdown to PDF", "make a PDF of this doc", "turn these notes into a PDF", or wants to share a local .md as a PDF. Handles the macOS wkhtmltopdf 0-page trap and preserves monospace/aligned blocks.
Turn a Markdown report into a clean, branded A4 PDF the president can read or print. Use whenever the user asks for a report/summary/proposal "as a PDF" or "to read properly". Works fully offline on Windows using pandoc + headless Chrome/Edge — no LaTeX or extra installs needed.
Repair or replace short text in scanned or image-only PDF pages while preserving the visual appearance. Use when Codex needs to edit a PDF that has no text layer, replace one Chinese character or a short phrase, cover old pixels, match font size, weight, baseline, gray level, blur, and verify the rendered result with screenshots.
URL to PDF and HTML to PDF. Convert any web page URL, or a raw HTML string, into a clean PDF file. Save an invoice or a receipt, archive an article, or print an HTML report to PDF. The page renders in a real hosted browser with JavaScript on, so the PDF matches the live page, and you set paper size, margins, orientation, and backgrounds. No local headless Chrome or wkhtmltopdf to install. The agent registers its own key and gets free credits right away, so the first PDF works with no signup. A person can confirm one email to add more free credits, and only successful renders cost anything.
当用户明确要求"下载文献全文"或"获取论文PDF"时使用。通过 DOI 号下载学术论文全文 PDF,支持 arXiv、Sci-Hub、Unpaywall、期刊官网等多源策略。⚠️ 不适用:用户只是想解析或处理已有的 PDF 文件(应使用 pdf skill)、只是想搜索论文信息而无需下载全文、没有提供 DOI/标题/BibTeX 任何标识符。
Compress PDF files to reduce size. Uses Ghostscript when available, falls back to pikepdf automatically. Trigger when the user says things like "compress this PDF", "PDF is too large", "reduce PDF size", or equivalent in any language.
Export HTML files to PDF using Playwright with mobile emulation and custom viewport settings. Use this skill when the user asks to convert HTML to PDF, export presentations to PDF, or generate PDF documents from web pages.
Edit PDF files locally without Canva API. Use this skill when user provides a PDF file and wants to modify content, extract text, merge/split pages, or update specific sections. Maintains original formatting, spacing, fonts, and layout. Does NOT require Canva. Output goes to the output/ folder. Use for: "edit this PDF", "change page 15", "update text in PDF".
Use when the user asks for the page count of an existing PDF, whether a PDF meets an exact/minimum/maximum page limit, or the numeric page-count result for another workflow. Triggers on "PDF pages", "page count", "몇 페이지", and "분량 제한 확인". NOT for creating, editing, rendering, or visually reviewing a PDF; use the PDF skill for those tasks.
ifly-pdf-image-ocr skill supporting both image OCR (AI-powered LLM OCR) and PDF document recognition. Use when user asks to OCR images, extract text from images/PDFs, convert PDF to Word/Markdown, or perform any OCR tasks on images or PDFs. Supports multi-language text extraction, document layout understanding, and various output formats.
Use when the user wants to extract text, tables, or images from a PDF file — including scanned/image-only PDFs needing OCR — or asks to "read this PDF", "extract text from PDF", "what's in this PDF", or to pull tables/images out of one.
Plain text files in a repository that tell a coding agent how the project works: commands to run, conventions to follow and things to avoid. CLAUDE.md, AGENTS.md, cursor rules and skills are the common kinds.
CLAUDE.md or AGENTS.md?
CLAUDE.md is read by Claude Code. AGENTS.md is an open format that Codex, Cursor and other agents read. Many projects keep one and point the other at it.
What is a skill?
A folder with a SKILL.md that describes one capability, such as filling PDFs or reviewing code. The agent loads it only when the task calls for it.
Can I search my own team's files too?
Your agents already can, over MCP, limited to the files you're allowed to read. Searching them from this page is coming.