with one click
vibecut
vibecut contains 10 collected skills from Nuva-Lab, with repository-level occupation coverage and site-owned skill detail pages.
Skills in this repository
Generate karaoke-style word-level timestamps by aligning script text to audio using Qwen3-ForcedAligner + jieba for Chinese word segmentation. Use when the user says 'align captions', 'karaoke timestamps', 'word timestamps', 'caption alignment', 'sync text to audio'.
Analyze raw video content using Gemini to identify speakers, topics, key moments, and potential clip opportunities
Audio processing utilities - noise reduction, normalization, enhancement
Extract a video segment using FFmpeg with precise start/end times
Find naturally clean, coherent video segments worth keeping (selection over repair)
Text-guided audio source separation using SAM-Audio via mlx-audio
ASR with ~30ms timestamp precision using Qwen3-ASR + ForcedAligner
Transcribe a video clip using Gemini to get timestamped segments for captions
Clone a voice using qwen3-tts and generate speech from text
Generate voiceover scripts in Joyce's style for video clips