Use this skill whenever the user wants to "拉片" (shot-by-shot analyze) a video, generate a director-friendly shot-breakdown report, extract keyframes from a video, transcribe its audio, detect scene cuts, write Seedance/Kling/Jimeng-style AIGC video reproduction prompts, or produce a polished HTML delivery page from a video file. Trigger this whenever the user mentions video shot analysis, scene detection, keyframe extraction, AIGC video prompt generation, or asks to turn a video into a structured analysis report (shots.json, frames, transcript, summary, HTML page). The skill handles the deterministic pipeline (probe → detect → extract → transcribe → align → build) via scripts; the visual analysis of each shot is performed by the Claude model running this conversation, using the extracted keyframes as multimodal input.
2026-04-27