Skip to main content

paddleocr-doc-parsing

Parse documents using PaddleOCR's API.

소스 정보

저장소
Kernel8901/ai-agent-skills-classification
최근 소스 활동
2026년 4월 4일 15:26
감지된 SKILL.md 언어
영어
스타
5
포크
1

설치 방법

기본적으로 소스를 먼저 확인하는 Prompt가 선택됩니다. 직접 명령으로 전환하거나 로컬 사본을 다운로드할 수도 있습니다.

소스 파일 검토

설치 여부를 결정하기 전에 SKILL.md와 SkillsMP에 표시된 보조 파일을 읽어 보세요.

파일 탐색기
3 개 파일

SKILL.md 표시 중

SKILL.md
소스 지침 · 읽기 전용 미리보기
name
paddleocr-doc-parsing
description
Parse documents using PaddleOCR's API.
homepage
https://www.paddleocr.com
metadata
{"openclaw":{"emoji":"📄","os":["darwin","linux"],"requires":{"bins":"[Truncated]","env":"[Truncated]"}}}
# PaddleOCR Document Parsing Parse images and PDF files using PaddleOCR's API. Supports multiple document parsing algorithms with structured output. ## Key Features - **Multi-format support**: PDF and image files (JPG, PNG, BMP, TIFF) - **Layout analysis**: Automatic detection of text blocks, tables, formulas - **Multi-language**: Support for 110+ languages - **Structured output**: Markdown format with preserved document structure ## Setup 1. Obtain credentials from the [PaddleOCR official website](https://www.paddleocr.com). Click the “API” button, choose the desired algorithm (e.g., PP-Structure, PaddleOCR-VL-1.5), and copy the API URL and the access token. 2. Set environment variables: ```bash export PADDLEOCR_API_URL="https://your-endpoint-here" export PADDLEOCR_ACCESS_TOKEN="your_access_token" ``` ## Usage Examples ### Run Script ```bash # Parse local image {baseDir}/paddleocr_parse.sh document.jpg # Parse local PDF file {baseDir}/paddleocr_parse.sh -t pdf document.pdf # Parse document from URL {baseDir}/paddleocr_parse.sh -t pdf https://example.com/document.pdf # Output to stdout (default) {baseDir}/paddleocr_parse.sh document.jpg # Save output to file {baseDir}/paddleocr_parse.sh -o result.json document.jpg ``` ### Response Structure ```json { "logId": "unique_request_id", "errorCode": 0, "errorMsg": "Success", "result": { "layoutParsingResults": [ { "prunedResult": [...], "markdown": { "text": "# Document Title\n\nParagraph content...", "images": {} }, "outputImages": [...], "inputImage": "http://input-image" } ], "dataInfo": {...} } } ``` **Important Fields:** - **`prunedResult`** - Contains detailed layout element information including positions, categories, etc. - **`markdown`** - Stores the document content converted to Markdown format with preserved structure and formatting. ## Quota Information See official documentation: https://ai.baidu.com/ai-doc/AISTUDIO/Xmjclapam
GitHub에서 보기