| name | storyboard-prompter |
| description | 把任何 idea 变成一条给图像生成模型(GPT-IMAGE-2 / Nano Banana / Seedream 等)的 storyboard prompt,一次性画出整张多格故事板(3×2 / 3×3 / 4×4 / 漫画页)。用户说"分镜"/"故事板 prompt"/"storyboard"/"分镜提示词"/"出整张分镜图"/"把这个场景排成分镜"/"idea to storyboard prompt"/"分镜脚本画面"时必用。自动把 idea 拆成每格 beat,钉死美术方向保证跨格一致性,每格带镜头语言(景别+运镜+时长)。支持 6 种风格切换(铅笔电影分镜/漫画/写实电影/水墨/吉卜力/赛博朋克)。输出一条可直接粘贴的英文 text prompt,不生成图片本身,不写代码,不调 API。 |
storyboard-prompter
把 idea 变成一条 prompt,喂给 GPT-IMAGE-2,一次性画出整张多格故事板。
这是什么 / 不是什么
- 是:idea → 1 条 text prompt → GPT-IMAGE-2 出整张多格故事板
- 不是:一镜一图的 N 条 prompt(那是另一个用法,本 skill 不做)
- 不是:自己调 API 生图(输出只是 prompt 文本,用户拿去喂模型)
- 不是:带转场/对白/音效的文字分镜脚本(那是导演脚本,不是图像 prompt)
工作流
- 收 idea:用户给一句话 idea(如"雨夜东京小巷追逐,最后跳屋顶")
- 三个变量(用户没说就用默认):
- 网格数:默认 6 格 3×2
- 风格:默认
pencil-storyboard(铅笔电影分镜)
- 比例:默认
landscape 16:9
- 自动拆 beat:按起承转合把 idea 拆成 N 个镜头,每格 1 个动作
- 填模板:每格 = 主体动作 + 镜头景别 + 运镜 + 时长 + info 条
- 钉美术:Art direction 钉死 palette / 角色 / 光照 / 风格锚点,保证 N 格像同一个故事
- 输出:一条 code block 的英文 prompt,用户复制粘贴喂 GPT-IMAGE-2
核心模板
A {N}-panel storyboard laid out as a {COLS}×{ROWS} grid, {ASPECT} overall.
Each panel is a rectangular sketch with a white margin border and a small info strip underneath.
(Style vocabulary — pencil / ink / film grain / etc. — belongs in Art direction below, NOT in this line. Stating it twice wastes the model's detail budget and contradicts itself when the two phrasings drift.)
Scene: {ONE-LINE IDEA}
Panel 1 — {SHOT-TYPE}: {ACTION/SUBJECT with 5-12 concrete nouns + explicit spatial prepositions for any technical action}. Info: "PANEL 1 · {INT/EXT} · {LOCATION} · {TIME} · {SHOT} / {CAMERA-MOVE} / {DURATION}{ / SFX: SOUND}"
Panel 2 — {SHOT-TYPE}: {ACTION/SUBJECT}. Info: "PANEL 2 · ..."
...
Panel {N} — {SHOT-TYPE}: {ACTION/SUBJECT}. Info: "PANEL {N} · ..."
Art direction: {STYLE-ANCHOR} — {PALETTE} — {LINE/RENDER} — {LIGHTING} — {CONSISTENCY-LOCK across panels}
Avoid: garbled text, inconsistent character appearance across panels, generic stunning/cinematic adjectives, fake brand logos, action-physics errors (wrong trajectory direction, body facing, tool contact point, joint angle), {style-specific drift}.
5 种网格变体
| 用途 | 网格 | 面板格式 | 何时用 |
|---|
| 电影分镜(默认) | 3×2 / 6 格 | 铅笔草图 + 白边 + info 条 | 用户说"分镜"/"storyboard"未指定类型 |
| 表情/姿态网格 | 4×4 / 16 格 | 同角色不同表情,无 info 条 | 用户说"表情"/"expression grid"/"角色不同状态" |
| 角色设定网格 | 8-10 格 | 前/侧/后视图 + 表情变体 + 部件拆解 | 用户说"角色设定"/"character sheet"/"三视图" |
| 漫画分镜页 | 多格编号缩略页 | 漫画页排版含对白框/速度线 | 用户说"漫画分镜"/"manga page"/"漫画页" |
| 世界构建大图 | 3×3 / 9 格 | 同一世界观下 9 个场景/角色 | 用户说"世界构建"/"worldbuilding"/"九宫格世界观" |
用户没指定时,默认 3×2 电影分镜。
风格锚点库
整段贴进模板的 Art direction 段。用户没指定就是 pencil-storyboard。
pencil-storyboard(默认)
classic animation-school storyboard — pencil line-work, grey marker shading, red-pencil arrow annotations on action and camera-move panels, off-white paper texture background
manga-panel
shōnen manga page — black ink line, screentone shading, speed lines on action panels, bold black panel borders, white gutters, dynamic panel shapes with diagonal cuts for impact
cinematic-photo
photoreal cinematic storyboard — 35mm film grain, natural lighting per scene, muted teal-orange grade, shallow depth of field, anamorphic lens flare, shot-on-film texture
ink-chinese
Chinese ink-wash painting (水墨) — monochrome ink with varied density, deliberate white space (留白), dry-brush flying-white strokes on motion (飞白), subtle color accent only on key subjects, rice-paper texture
ghibli-cel
Studio Ghibli hand-painted cel — watercolor-gouache background, thin ink line on characters, desaturated greens and warm skin tones, visible brush texture, soft atmospheric perspective, Miyazaki aesthetic, not 3D render
cyberpunk-neon
cyberpunk storyboard — neon-soaked rain, chrome reflections, magenta-cyan palette, high contrast, holographic UI overlays, blade-runner vibe, original characters only, no real brand logos
镜头语言词汇表
景别(Shot type):
WIDE / EWS(全景/大远景)· MEDIUM / MS(中景)· OTS(过肩)· CU(特写)· ECU(大特写)· low angle(仰拍)· high angle(俯拍)· aerial(航拍)· POV(主观视角)· match cut(匹配剪辑)
运镜(Camera move):
static(固定)· pan-L / pan-R(左右摇)· tilt-up / tilt-down(上下俯仰)· dolly-in / dolly-out(推拉)· crane-up / crane-down(升降)· follow-cam(跟随)· handheld(手持晃动)· 360° orbit(环绕)
时长:1s / 1.5s / 2s / 3s / 4s(默认每格 2s)
info 条格式:
PANEL N · INT/EXT · LOCATION · TIME · SHOT-TYPE / camera-move / duration / SFX: sound
自动拆 beat 规则
收到一句 idea 后按下面规则拆成 N 格:
- 默认 6 格 3×2,叙事弧:建立 → 引入 → 发展 → 冲突 → 高潮 → 落版
- 每格 1 个动作:一格一动作一镜头。不要 "Panel 1:他跑到巷口然后被追上然后反抗"——拆成 3 格
- 景别递进:通常 WIDE 开场建立场景 → 中段切 CU/OTS 拍情绪和冲突 → 末尾回 WIDE 或 match cut 落版
- 场景密度:每格 5-12 个具体名词(不是形容词)。例:
wet neon alleyway, kanji signage, discarded umbrella, puddle reflection, steam vent
- 角色一致性:idea 里有反复出现的角色时,第一格建立其外观(服装/颜色/特征),后续格只引用不重述,Art direction 段钉死身份
- 用户指定 N 时:按 N 拆,叙事弧压缩或拉伸到 N 格(4 格=建立/发展/高潮/落版;9 格=三幕每幕 3 格)
动作 / 物理逻辑校验(最容易翻车,必做)
idea 含冷门 / 专业 / 反直觉 / 有特定空间几何的技术动作时(体育技术、武术招式、工艺操作、乐器演奏、舞蹈、急救手法、工业流程、烹饪手法……),拆 beat 前先做物理逻辑校验。常见日常动作不校验(走路 / 跑步 / 坐下 / 喝水 / 挥手打招呼 / 拥抱 / 上下楼梯)——这些模型画得对,堆 FROM BEHIND / IN FRONT 反而啰嗦。
为什么必做:GPT-IMAGE-2 对常见动作有强先验,会按"最常见类似动作"默认画。冷门或专业动作的默认画法几乎必错——例如彩虹过人会被画成球从身前弧线飞过(实际是从身后夹起、越头顶、落身前)。只写动作名等于没写。
流程
- 识别动作:idea 里有没有高风险技术动作?是哪一个?(判断标准见下方"高风险动作的识别特征")
- 找模型容易画错的点:这个动作的关键空间关系(轨迹方向 / 相对位置 / 接触点 / 关节角度 / 平面 / 时间序列)是什么?模型默认会画成什么样?
- 用显式约束标注关键空间关系:把模型容易画错的点写进对应 panel 的描述。只写动作名等于没写。约束手段按动作类型选,方位介词只是默认手段之一:
- 相对位置类(彩虹过人球路)→ 方位介词:
ball arcs FROM BEHIND the attacker, OVER both heads, lands IN FRONT
- 轨迹形状类(弧线 vs 直线)→ 显式描述:
ball travels in a vertical arc, NOT a straight horizontal pass
- 接触点类(焊接 / 乐器握姿)→ 指明接触点:
arc strikes AT the contact point between electrode and base metal
- 关节角度类(钢琴 / 投篮)→ 指明角度:
fingertips on keys, knuckles raised, wrist below hand-back
- 时间序列类(多步动作)→ 标序:
step 1 first (heel trap), THEN step 2 (flick up), THEN step 3 (ball arcs over)
某动作反复画错时,追加 negative prompt:Avoid: [常见错法]。
- 默认靠物理常识 + 显式约束;只在关键几何拿不准时才停下问:80% 情况 Claude 自己懂动作的物理(彩虹过人球走身后、跳高背越式头先过杆),直接用规则 3 的约束写显式即可。web search / 问用户是兜底,不是默认——只有在 Claude 自己都对关键几何犯嘀咕时(比如"投飞镖手腕到底锁不锁"),才在输出 prompt 之前问用户一句,不要把不确定的物理写死。瞎编的 prompt 会让模型画出更自信的错误。
高风险动作的识别特征(清单只是样本,不是覆盖表)
判断一个新动作要不要校验,看它有没有以下特征之一,而不是查它在不在这张表上:
- 冷门(训练数据少,模型先验弱):彩虹过人、背越式跳高、特定武术招式、地方技艺
- 反直觉(默认画法和实际相反):球路 / 轨迹方向、身体朝向、过杆姿势
- 有特定空间几何(接触点 / 平面 / 角度独特):焊接引弧点、挥棒平面、握姿
- 多步时间序列(顺序错了就全错):投掷发力链、舞蹈连接步
craft 硬规则(每条都带 why)
1. 先说结构,再说画面
模板开头必须先写 A 6-panel storyboard laid out as a 3×2 grid, landscape 16:9。
Why:GPT-IMAGE-2 把细节预算花在最先读到的东西上。先说网格,模型按布局分配预算;先说场景,模型把预算全花在主体上,布局会 improvisation。
2. 每格一个 beat
一格一动作一镜头。不要塞多个动作。
Why:一格塞多个动作,模型把它们全画进同一格,画面糊成一团。
3. 共享美术锁一致性
所有格共享 palette、角色身份、服装、光照、风格锚点。钉在 Art direction 段,不在每格重复。
Why:GPT-IMAGE-2 画多格时每格是独立采样。没有共享锚点,6 格像 6 个不同故事。
4. 风格锚点要具体有界
写 MAPPA-style digital 2D animation 或 classic animation-school pencil storyboard,不写 professional / cinematic / beautiful。
Why:空形容词不指向任何具体视觉参考,模型自由发挥 = 不可预测。具体锚点 = 模型能匹配训练分布。
5. 场景密度 > 形容词
每格 5-12 个具体名词(kanji signage / wet cobblestones / copper whisk),不堆 stunning/cinematic。
Why:名词可视化,形容词不可视。"雨夜小巷"模糊,"wet neon alleyway, kanji signage, discarded umbrella, steam vent"精确。
6. 镜头语言必带
每格 info 条必须含景别 + 运镜 + 时长。
Why:故事板区别于"多格插画"的关键。镜头语言让它读起来像分镜,不像九宫格装饰画。
7. avoid 清单短而准
默认加一行:Avoid: garbled text, inconsistent character appearance across panels, generic stunning/cinematic adjectives, fake brand logos.
Why:模型有常见坏默认(文字乱码、角色 drift、堆空形容词)。短 avoid = 强先验;长 avoid = 主导整个 prompt 反而坏事。上限 5 条。
输出格式
输出一条 code block 的英文 prompt,用户复制粘贴喂 GPT-IMAGE-2。
不要:
- 输出对 prompt 的解释(除非用户问)
- 输出多条变体(除非用户要)
- 输出中文 prompt(GPT-IMAGE-2 英文效果更稳)
- 自己调 API 生图(用户自己喂)
如果 idea 太模糊("画个有趣的故事板"),只问 1 个最关键的问题(主角是谁 / 场景在哪 / 想要什么情绪),不要问一长串。
示例
Input:雨夜东京小巷追逐,最后跳屋顶,要水墨风
Output:
A 6-panel storyboard laid out as a 3×2 grid, landscape 16:9 overall. Each panel is a rectangular sketch with a white margin border and a small brush-calligraphy info strip underneath.
Scene: a chase through a rainy alleyway ending in a rooftop leap, ink-wash wuxia tone.
Panel 1 — WIDE establishing: lone runner entering from left, narrow alley walls with hanging lanterns, rain streaks as vertical ink dashes, puddle reflections. Info: "PANEL 1 · EXT. ALLEY · NIGHT · WIDE / static / 2s / SFX: rain"
Panel 2 — OTS tracking: runner mid-stride from behind, pursuer silhouette 10 m back emerging from mist, torn banner fluttering. Info: "PANEL 2 · OTS TRACKING / follow-cam / pan-L 45° / 3s"
Panel 3 — CU: runner's face half-turned, eyes darting up toward fire escape, sweat as ink droplets, hair plastered. Info: "PANEL 3 · CU RUNNER / static / 1.5s / SFX: breath"
Panel 4 — Low angle: runner leaping onto fire-escape ladder, robe flaring as flying-white stroke, ladder rungs rusted. Info: "PANEL 4 · LOW ANGLE / tilt-up 30° / 2s"
Panel 5 — Wide aerial: runner silhouetted against tile-roof skyline about to leap between buildings, mist swallowing the pursuer below, broken moon. Info: "PANEL 5 · WIDE AERIAL / crane-down / 4s"
Panel 6 — Match cut: runner's straw sandals landing on wet rooftop, water splash as ink burst, scattered cherry petals. Info: "PANEL 6 · MATCH CUT CU / static / 1s / SFX: splash"
Art direction: Chinese ink-wash painting (水墨) — monochrome ink with varied density, deliberate white space (留白), dry-brush flying-white strokes on motion (飞白), subtle vermillion accent only on the runner's sash for identity lock, rice-paper texture. Shared across all 6 panels: same runner (straw sandals, dark robe, vermillion sash), same lantern-lit alley palette, same rain-as-ink-dash rendering.
Avoid: garbled calligraphy, inconsistent runner appearance across panels, color clutter, manga/anime style drift.