用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
直接命令不会经过审查 Prompt;运行前请先检查来源。
npx skills add https://github.com/LYL1015/JarvisHub --skill flova-dialogue-interview-video命令会保持在同一行。复制前请横向滚动并检查完整内容。
想先保存到本地?可下载 SkillsMP 当前能够提供的文件。
Use when the user asks JarvisHub AI Chat or a `webHero` workflow to generate a website, landing page, homepage, brand-style web design, preview-first webpage workflow, or wants section screenshots, webpage assets, and final HTML to be produced as one staged flow.
把 assets/demo 提炼成运行时可用的视觉连续性方法论:先锁资产/角色/场景/镜头语义,再做扩镜与视频。(素材生成场景:角色一致性、视觉锚点锁定、单素材生成前的锚点规范与一次只改一个变量。)
基于用户提供的设计稿图片还原真实网页。该 skill 可用于任意前端项目,但强依赖本地已启动且已授权的 JarvisHub vision 工作流;视觉模型由服务端默认配置决定,支持外部传入 vision prompt,且不可静默降级。
基于 SOC 职业分类
正在显示 SKILL.md
| name | flova-dialogue-interview-video |
| description | Use when 用户要制作多人对话、访谈、采访、圆桌、谈话节目、角色正反打或音频驱动口型/表演视频,并需要发言镜头、反应镜头、音轨和时间线控制。 |
用于多人对话与访谈视频:为采访、圆桌访谈和角色谈话场景设计 key elements、对话音频层、关键帧、发言镜头、反应镜头和最终时间线。
This workflow is dialogue-first. Audio and speaker turns control the edit.
imageUrl / videoUrl,不能把提交态当完成态。critic sub-agent:只读取真实媒体并评审,不生成、不补素材。blocked 项,除非本轮工具列表明确暴露对应能力。Use this skill for:
Use flova-scripted-short-production for broader drama scenes with action-heavy story structure. Use this skill when the problem is specifically dialogue performance and conversational editing.
Follow the JarvisHub Execution Model. Use audio-driven dialogue video only when this turn exposes a real audio-driven tool.
If this turn exposes audio-driven image-to-video, use it for dialogue shots. If not, generate visual performance clips and preserve dialogue as script/timeline metadata or bind user-provided audio assets.
Pause after:
Collect:
If dialogue is not scripted, ask whether to generate an interview script or use bullet points.
For each speaker:
For the set/location:
Provided assets must be reused and bound before generating substitutes.
Each shot should be one of:
speaker_single: active speaker close/medium close.two_shot: two speakers in one frame.over_the_shoulder: foreground listener, focus on speaker.reaction: listener reaction while another speaker talks.host_or_moderator: question/transition.cutaway: prop, audience, notes, environment.For alternating dialogue, split into separate shots if two people speak back and forth. Do not pack many alternating speakers into one shot unless the current tool can reliably handle it.
Shot fields:
shot_id,speaker,dialogue_text,audio_asset_or_plan,shot_type,framing,camera_angle,listener_reaction,body_action,keyframe_goal,duration_target,timeline_position.Use reaction shots to prevent flat talking-head sequences.
Audio layers are the master when real audio exists; otherwise the dialogue script and timing plan are the master:
If using audio-driven video:
If not using audio-driven video:
BGM should stay low under speech or be omitted.
For dialogue keyframes:
For reaction keyframes:
Generate later keyframes in the same setting using earlier similar keyframes as continuity references when supported.
Dialogue shot prompt should describe dynamic change, not re-describe the static keyframe.
Template for audio-driven dialogue shot:
Camera [movement or hold] on [speaker]. [Speaker] speaks the attached audio with [emotional state], [face/body action], while [listener/background action]. Keep the speaker identity and eyeline from the keyframe. no music, no subtitles, no random text, no watermark.
Template for non-dialogue reaction shot:
[Listener] remains silent, [specific micro-expression and body reaction], eyes shift toward [speaker/object], [small camera behavior]. no dialogue, no music, no subtitles, no random text, no watermark.
Avoid:
Audio controls pacing:
Use visual variation:
Check:
If lip sync is poor, reduce reliance on active-mouth closeups and use reaction/OTS/cutaway edits.
Return:
spec,participants,dialogue_or_interview_outline,shot_list,audio_layers,keyframes,video_prompts,media,assembly_plan,review.Do not present audio-less keyframes as a finished interview video.