con un clic
gemini-vibecut
gemini-vibecut contiene 11 skills recopiladas de Nuva-Lab, con cobertura ocupacional por repositorio y páginas de detalle dentro del sitio.
Skills en este repositorio
Align text to audio timestamps using Qwen3-ForcedAligner (~30ms precision)
Full pipeline from manga panels to animated video. Two audio modes: dialogue (Qwen3-TTS + karaoke captions) or music (ElevenLabs + rolling lyrics).
Burn karaoke captions into video using FFmpeg ASS subtitles (~20s for 16s video)
Generate multi-panel manga from character reference and story beats
Generates original background music using Google's Music Generation API. Creates soundtracks matched to scene mood and timing.
Concatenates video clips and optionally adds background music using FFmpeg.
Generates anime character images from photo analysis using Nano Banana Pro (gemini-3-pro-image-preview). Creates character sheets with full body and portrait views at 2K resolution.
Generates speech audio from dialogue using Gemini TTS. Supports multi-speaker conversations with consistent voices per character.
Generates animated video using Veo 3.1 multi-image reference mode with native audio. No TTS needed.
Text-to-speech generation using Qwen3-TTS via FAL API
Analyzes photos to extract structured information about pets, people, or world settings using Gemini 3 multimodal capabilities.