用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
直接命令不会经过审查 Prompt;运行前请先检查来源。
npx skills add https://github.com/tigerowo/infinite-canvas --skill h3-skill命令会保持在同一行。复制前请横向滚动并检查完整内容。
想先保存到本地?可下载 SkillsMP 当前能够提供的文件。
正在显示 SKILL.md
| name | H3 官方默认 Skill |
| description | 为 MiniMax H3 的五种视频生成模式(T2VA、I2VA、FL2VA、L2VA 和 Ref2VA)编写结构化的提示信息 |
| compatibility | Portable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only adds optional ChatGPT/Codex UI metadata; it does not restrict the skill to OpenAI agents. |
references/base-en.txt and follow its final prompt structure.references/ref-en.txt for the remaining reference-specific rules and complete example.Use integrated_multimodal_description, overall_soundscape, and non_diegetic_music in the order shown in references/base-en.txt.
Ref2VA rewrites use subject_definitions, summary, retention_analysis, detailed_description, overall_soundscape, and non_diegetic_music in that order. Reference labels stay consistent across all sections.
This guide explains how rewrite outputs are organized and written in full-reference mode.
Write all six rewrite sections in English. Preserve the original language only for dialogue and lyrics inside <d> and for text visibly present in the scene.
Description detail: Make detailed_description as detailed and explicit as possible. For each shot, clearly establish the current composition, subject appearance and position, environment and lighting, actions and state changes, camera movement, current sound, and the points where referenced content actually appears or takes effect. Avoid reducing the description to a plot summary or a list of reference relationships.
The basic formats for shots, camera movement, speakers, dialogue, and ordinary sound are shared with the Video Prompt Writing Guide (T2VA / I2VA / FL2VA / L2VA). This guide focuses on the reference labels, analysis sections, and format differences specific to full-reference mode.
A complete rewrite output consists of six sections in the following order:
| Section | Purpose |
|---|---|
subject_definitions | Defines referenced content and its reference labels |
summary | Summarizes the task type, target video, and main reference relationships |
retention_analysis | Describes how referenced content is preserved, transferred, or reused |
detailed_description | Describes visuals, actions, shots, sound, and dialogue in playback order |
overall_soundscape | Summarizes ambience and physical sounds |
non_diegetic_music | Describes background music audible only to the audience |
The basic format follows the Video Prompt Writing Guide (T2VA / I2VA / FL2VA / L2VA):
[Shot 1] marks the opening shot and has no timestamp. Later shots use [Shot N] At MM:SS.mmm, ... to mark cut times.(S1), (S2), and subsequent IDs. Write dialogue and lyrics as <d>[Language] ...</d>.<scenetrans>, <cutoff>, and the corresponding continuity descriptions for dialogue crossing a cut, speech truncated by the video ending, and continuous audio across shots.For complete rules and examples covering camera vocabulary, group speech, voice-over, dialogue across cuts, and visible text, see the Video Prompt Writing Guide (T2VA / I2VA / FL2VA / L2VA).
overall_soundscape and non_diegetic_musicThe definitions of these two sound categories follow the Video Prompt Writing Guide (T2VA / I2VA / FL2VA / L2VA).
overall_soundscape summarizes ambience and physical sounds across the full video. Dialogue, singing, and sound events synchronized to a particular shot remain in detailed_description:
overall_soundscape: Quiet indoor room tone and a low ventilation hum continue throughout the video.
non_diegetic_music describes background music that the characters cannot hear and that is audible only to the audience. When music is present, state its instrumentation, tempo, and dynamic development:
non_diegetic_music: A restrained solo-piano score at a slow tempo, with sustained low cello underneath and no swell.
When reference audio is used, state its copy or reference relationship only in the section that matches the audible layer: ambience and sound effects belong in overall_soundscape, while audience-only score belongs in non_diegetic_music. If the same audio provides both kinds of content, describe the corresponding relationship in each section:
overall_soundscape: The copied ambience layer from <Audio 1> continues throughout the target video.
non_diegetic_music: <Audio 2> is directly reused as the complete audience-only score.
Write complete dialogue and lyrics only inside <d> in detailed_description; do not repeat them in these two sections.
<Picture 1>, <Video 1>, <Audio 1>) across every section.