Skip to main content

mineru-pdf

Stars8
Forks1
UpdatedJune 4, 2026 at 16:49

High-accuracy PDF content extraction using MinerU (Shanghai AI Lab). Use this whenever the user needs to extract text, formulas, tables, or images from a complex PDF — especially academic papers, multi-column layouts, scanned documents, or any PDF where pypdf produces garbled /Cxx formula output. Trigger on: "MinerU", "解析这篇论文", "提取PDF公式", "这个PDF乱码了", "扫描件OCR", "双栏PDF", or any request involving PDF formula/image extraction. Even if the user doesn't say "MinerU" explicitly, suggest it when they complain about garbled PDF text or need high-quality extraction. For simple single-column text-only PDFs, the default pdf skill is faster.

Installation

Install with Codex or Claude Copy this prompt, paste it into Codex, Claude, or another assistant, and let it review the skill page and install it for you.

SKILL.md
readonly