High-accuracy PDF content extraction using MinerU (Shanghai AI Lab). Use this whenever the user needs to extract text, formulas, tables, or images from a complex PDF — especially academic papers, multi-column layouts, scanned documents, or any PDF where pypdf produces garbled /Cxx formula output. Trigger on: "MinerU", "解析这篇论文", "提取PDF公式", "这个PDF乱码了", "扫描件OCR", "双栏PDF", or any request involving PDF formula/image extraction. Even if the user doesn't say "MinerU" explicitly, suggest it when they complain about garbled PDF text or need high-quality extraction. For simple single-column text-only PDFs, the default pdf skill is faster.
使用 MinerU(上海 AI Lab)进行高精度 PDF 内容提取。 当用户需要从复杂 PDF 中提取文本、公式、表格或图片时使用此 skill — 尤其是学术论文、多栏排版、扫描文档,或任何 pypdf 产生 /Cxx 乱码的 PDF。 触发词:"MinerU"、"解析这篇论文"、"提取 PDF 公式"、"这个 PDF 乱码了"、 "扫描件 OCR"、"双栏 PDF",或任何涉及 PDF 公式 / 图片提取的请求。 即使用户没有明确说 "MinerU",当他们抱怨 PDF 文本乱码或需要高质量提取时, 主动建议使用此 skill。对于简单的单栏纯文字 PDF,默认 pdf skill 更快。