Skip to main content

infinity-parser2

Parse PDFs, scanned documents, and document images (screenshots, invoices, paper documents, whiteboard photos) into structured layout JSON (labeled regions with bounding boxes) and Markdown via Infinity-Parser2 through an existing vLLM server. Use when the user wants to extract text, tables, formulas, or structured data from visual documents; mentions OCR, text recognition, or document parsing; or asks to parse, digitize, or extract content from a document. Requires a vLLM OpenAI-compatible API URL and API key.

インストールへ移動

ソース情報

リポジトリ
infly-ai/INF-MLLM
ソースの最終更新活動
2026年8月31日 10:05
検出された SKILL.md の言語
英語
スター
249
フォーク
27

インストール方法

デフォルトでは、最初にソースを確認する Prompt が選択されています。直接コマンドに切り替えるか、ローカルコピーをダウンロードすることもできます。

ソースファイルを確認

インストールを決める前に、SKILL.md と SkillsMP に表示されている付属ファイルをお読みください。