Skip to main content

doc-ocr

文档文字识别。用户提供 PDF/扫描件/图片(合同、发票、书页、截图),需要提取文字、转成可编辑文本时使用。扫描件自动 OCR(macOS Vision,中英文)。Document OCR: extract editable text from PDFs, scans, and images (contracts, invoices, book pages, screenshots) via macOS Vision.

Aller à l'installation

Informations de source

Dépôt
davepoon/buildwithclaude
Dernière activité de la source
21 août 2026 à 06:58
Langue détectée de SKILL.md
chinois
Étoiles
3 512
Forks
520

Options d'installation

Le prompt qui vérifie d'abord la source est sélectionné par défaut. Vous pouvez passer à une commande directe ou télécharger une copie locale.

Vérifiez les fichiers source

Lisez SKILL.md et les fichiers associés affichés par SkillsMP avant de décider de l'installer.

Explorateur de fichiers
2 fichiers

Affichage de SKILL.md

SKILL.md
Instructions source · Aperçu en lecture seule
name
doc-ocr
description
文档文字识别。用户提供 PDF/扫描件/图片(合同、发票、书页、截图),需要提取文字、转成可编辑文本时使用。扫描件自动 OCR(macOS Vision,中英文)。Document OCR: extract editable text from PDFs, scans, and images (contracts, invoices, book pages, screenshots) via macOS Vision.
category
document-processing
license
MIT
# Doc-OCR 文档文字识别 PDF / 扫描件 / 图片 → 可编辑文字。有文字层的 PDF 直接提取,扫描件自动 OCR(macOS Vision 自带,中英文)。 ## 触发条件 用户提供 PDF/图片文件,要求: - "提取文字""转文字""OCR" - 处理扫描件、合同、发票、书页、截图 ## 使用步骤 ### 1. 单个文件 ```bash python3 scripts/dococr.py 合同.pdf python3 scripts/dococr.py 发票.jpg ``` 输出保存为 `<输入名>_ocr.txt`。 ### 2. 批量目录 ```bash python3 scripts/dococr.py ./扫描件/ -o 全部.txt ``` ### 3. Markdown 输出 ```bash python3 scripts/dococr.py 书.pdf --md ``` ## 依赖(首次使用时安装) ```bash pip3 install pymupdf pyobjc-framework-Vision ``` **注意**:OCR 依赖 macOS Vision(仅 macOS 可用)。Linux 需另装 tesseract 等引擎。 ## 已知陷阱 - **扫描件判定**:PDF 文字层 <20 字自动走 OCR,正常 PDF 直接提取。 - **手写体**:Vision 对印刷体/清晰手写效果好,潦草手写不保证。 - **隐私卖点**:文件在本机处理,不上传第三方。
Voir sur GitHub