Skip to main content
在 Manus 中运行任何 Skill
一键导入
GitHub 仓库

gemini-media-mcp

gemini-media-mcp 收录了来自 mordor-forge 的 4 个 skills,并提供仓库级职业覆盖和站内 skill 详情页。

已收集 skills
4
Stars
7
更新
2026-04-02
Forks
4
职业覆盖
3 个职业分类 · 已分类 100%
仓库浏览

这个仓库中的 skills

gemini-image-gen
平面设计师

Interactive AI image generation, editing, and composition using the gemini-media MCP (Google Gemini image models). Use this skill whenever the user asks to generate, create, draw, design, or make an image, illustration, photo, artwork, mockup, logo, icon, sticker, or any visual content. Also use when the user provides a description and wants it turned into a picture, asks to edit or modify an existing image, wants to compose multiple reference images, or mentions "gemini image", "image generation", "generate an image", or similar. This skill handles the full workflow from understanding intent through prompt engineering to model selection and iterative refinement.

2026-04-02
music-gen
录音工程技术员

Interactive AI music generation using the gemini-media MCP (Google Lyria 3 models). Use this skill whenever the user asks to generate, create, compose, or make music, a song, a beat, a soundtrack, a jingle, background music, or any audio music content. Also use when the user wants to create a melody, instrumental track, song with vocals, podcast intro music, or describes music they want to hear. Triggers on "make me a song", "generate music", "create a beat", "compose a soundtrack", "I need background music for...", "make a jingle", or any music creation request. This skill handles the full workflow from understanding musical intent through prompt construction with structure tags and lyrics to model selection and iterative refinement.

2026-04-02
tts-gen
录音工程技术员

Interactive text-to-speech audio generation using the gemini-media MCP (Google Gemini TTS). Use this skill whenever the user asks to convert text to speech, generate spoken audio, create a voiceover, narrate text, read something aloud, or produce audio from written content. Also use when the user wants a specific voice or language for audio output, mentions "text to speech", "TTS", "voiceover", "narration", "read this aloud", "speak this", or wants audio versions of text content. This skill handles voice selection, language configuration, and the generation workflow.

2026-04-02
video-gen
影片与视频编辑

Interactive AI video generation using the gemini-media MCP (Google Veo 3.1 models). Use this skill whenever the user asks to generate, create, or make a video, clip, animation, or motion content. Also use when the user wants to animate an existing image into video, extend a video clip, create a short film, promotional video, or any moving visual content. Triggers on "generate a video", "make a clip", "animate this image", "create a video of...", "video generation", or similar requests. This skill handles the full workflow from understanding intent through prompt engineering to async generation management and iterative refinement.

2026-04-02