Skip to main content
Jeden Skill in Manus ausführen
mit einem Klick

multimodal-structured-extraction

Sterne0
Forks0
Aktualisiert23. Mai 2026 um 12:34

Use this skill when extracting structured data from images (receipts, menus, statements, IDs, business cards, forms, screenshots) using a vision-capable LLM. Covers schema-first design with zod, validation gates and retry, locale handling (currency / dates / romanization), confidence signalling, and the anti-patterns that cause silent garbage extraction. Triggers on "extract from image", "OCR with structure", "vision model JSON", "Gemini multimodal", "GPT-4 vision extraction", "parse receipt", "parse business card", "image to schema", "structured output from image".

Installation

Mit Codex oder Claude installieren Kopieren Sie diesen Prompt, fügen Sie ihn in Codex, Claude oder einen anderen Assistant ein und lassen Sie die Skill-Seite prüfen und installieren.

SKILL.md
readonly