con un clic
mixing
mixing contiene 5 skills recopiladas de thorwhalen, con cobertura ocupacional por repositorio y páginas de detalle dentro del sitio.
Skills en este repositorio
Use for AUDIO editing with the `mixing` package: trim/crop a clip, fade in or fade out, join/concatenate clips (with optional crossfade), mix music under a voice or overlay one track on another, normalize loudness, convert to mono, resample, peek at raw samples, align two recordings of the same performance, or split a long recording (concert, DJ mix, radio show, podcast) into segments. Trigger on phrasings like "trim audio", "fade out the end", "join these audio clips", "add background music under the narration", "make it louder / normalize", "convert to mono", "align two recordings", "find the offset between these takes", "split a concert into songs", "separate speech from music". For video audio (replace/normalize a video's track) use mixing-video; this skill is audio-file → audio-file.
Use when re-voicing or translating audio/video with text-to-speech via the `mixing` package: turning text into spoken audio, synthesizing narration, translating subtitles, or producing a foreign-language dub of a video from its SRT. Trigger on requests like "dub this video", "re-voice this in another language", "text to speech / TTS this", "synthesize narration", "translate these subtitles", "make a French version of this clip", or "read this script in voice X". For transcribing speech → SRT in the first place, see mixing-transcript; for the `output` protocol and ffmpeg, see the mixing router skill.
Use the `mixing` Python package for video and audio editing tasks: slicing audio/video, fades, cropping, looping, changing speed, replacing or mixing a video's audio, normalizing levels, Ken Burns pan/zoom, concatenating clips, thumbnails, burning in subtitles, speech-to-text + filler removal, chapter detection, and TTS dubbing/translation. Trigger whenever someone wants to edit, transform, transcribe, dub, or assemble audio/video files with Python — e.g. "trim this clip", "add music to a video", "remove the ums", "make a thumbnail", "turn this audio into a podcast", "dub this in French". Start here to pick the right tool and the right sub-skill (mixing-audio / mixing-video / mixing-transcript / mixing-dubbing).
Transcribe speech to text, remove filler words (ums/uhs), build SRT subtitles or clean prose, and detect chapter markers — using the `mixing` package's `transcript` and `chapters` modules (ElevenLabs Scribe under the hood). Trigger whenever someone wants to "transcribe this audio/video", "get an SRT/captions", "remove the ums and uhs", "clean up the fillers", "make chapters/timestamps", "detect topic shifts", or "get word-level timestamps" from a media file. For re-voicing/translation use mixing-dubbing; for plain audio/video edits use mixing-audio / mixing-video. Read the `mixing` router skill first for the `output` egress protocol and ffmpeg/key setup.
Use for VIDEO editing with the `mixing` Python package: trim/crop a clip, loop it, speed it up or down, replace or mix in a new audio track, normalize audio levels, pan/zoom a still image (Ken Burns), build a film from panels, concatenate clips, grab/extract frames, make a YouTube-style thumbnail, burn in subtitles, or resize for shorts/tiktok/square. Trigger on phrasings like "trim this video", "loop the intro", "speed up / slow down this clip", "replace the audio with this music", "add a pan-zoom over this photo", "make a thumbnail", "burn the subtitles in", "resize for YouTube Shorts", or "stitch these clips together". For audio-only work use mixing-audio; for transcription / filler removal use mixing-transcript; for TTS re-voicing use mixing-dubbing.