Skip to main content
Manus에서 모든 스킬 실행
원클릭으로
GitHub 저장소

gemini-media-mcp

gemini-media-mcp에는 mordor-forge에서 수집한 skills 4개가 있으며, 저장소 수준 직업 범위와 사이트 내 skill 상세 페이지를 제공합니다.

수집된 skills
4
Stars
7
업데이트
2026-04-02
Forks
4
직업 범위
직업 카테고리 3개 · 100% 분류됨
저장소 탐색

이 저장소의 skills

gemini-image-gen
그래픽 디자이너

Interactive AI image generation, editing, and composition using the gemini-media MCP (Google Gemini image models). Use this skill whenever the user asks to generate, create, draw, design, or make an image, illustration, photo, artwork, mockup, logo, icon, sticker, or any visual content. Also use when the user provides a description and wants it turned into a picture, asks to edit or modify an existing image, wants to compose multiple reference images, or mentions "gemini image", "image generation", "generate an image", or similar. This skill handles the full workflow from understanding intent through prompt engineering to model selection and iterative refinement.

2026-04-02
music-gen
음향 기술자

Interactive AI music generation using the gemini-media MCP (Google Lyria 3 models). Use this skill whenever the user asks to generate, create, compose, or make music, a song, a beat, a soundtrack, a jingle, background music, or any audio music content. Also use when the user wants to create a melody, instrumental track, song with vocals, podcast intro music, or describes music they want to hear. Triggers on "make me a song", "generate music", "create a beat", "compose a soundtrack", "I need background music for...", "make a jingle", or any music creation request. This skill handles the full workflow from understanding musical intent through prompt construction with structure tags and lyrics to model selection and iterative refinement.

2026-04-02
tts-gen
음향 기술자

Interactive text-to-speech audio generation using the gemini-media MCP (Google Gemini TTS). Use this skill whenever the user asks to convert text to speech, generate spoken audio, create a voiceover, narrate text, read something aloud, or produce audio from written content. Also use when the user wants a specific voice or language for audio output, mentions "text to speech", "TTS", "voiceover", "narration", "read this aloud", "speak this", or wants audio versions of text content. This skill handles voice selection, language configuration, and the generation workflow.

2026-04-02
video-gen
영화·비디오 편집자

Interactive AI video generation using the gemini-media MCP (Google Veo 3.1 models). Use this skill whenever the user asks to generate, create, or make a video, clip, animation, or motion content. Also use when the user wants to animate an existing image into video, extend a video clip, create a short film, promotional video, or any moving visual content. Triggers on "generate a video", "make a clip", "animate this image", "create a video of...", "video generation", or similar requests. This skill handles the full workflow from understanding intent through prompt engineering to async generation management and iterative refinement.

2026-04-02