Skip to main content

audio

Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.

설치로 이동

소스 정보

저장소
Jinhong270/AI-Agent
최근 소스 활동
2026년 6월 29일 12:28
감지된 SKILL.md 언어
영어
스타
2
포크
0

설치 방법

기본적으로 소스를 먼저 확인하는 Prompt가 선택됩니다. 직접 명령으로 전환하거나 로컬 사본을 다운로드할 수도 있습니다.

소스 파일 검토

설치 여부를 결정하기 전에 SKILL.md와 SkillsMP에 표시된 보조 파일을 읽어 보세요.

파일 탐색기
6 개 파일

SKILL.md 표시 중

SKILL.md
소스 지침 · 읽기 전용 미리보기
name
Audio
slug
audio
version
1.0.1
description
Process, enhance, and convert audio files with noise removal, normalization, format conversion, transcription, and podcast workflows.
changelog
Declare required binaries (ffmpeg, ffprobe), add requirements section with optional deps, add explicit scope
metadata
{"clawdbot":{"emoji":"🔊","requires":{"bins":"[Truncated]"},"os":["linux","darwin","win32"]}}
## Requirements **Required:** - `ffmpeg` / `ffprobe` — core audio processing **Optional (for advanced features):** - `sox` — additional noise reduction - `whisper` — local transcription (or use API) - `demucs` — stem separation ## Quick Reference | Situation | Load | |-----------|------| | FFmpeg commands by task | `commands.md` | | Loudness standards by platform | `loudness.md` | | Podcast production workflow | `podcast.md` | | Transcription workflow | `transcription.md` | ## Core Capabilities | Task | Method | |------|--------| | Convert formats | FFmpeg (`-acodec`) | | Remove noise | FFmpeg filters or SoX | | Normalize loudness | `ffmpeg-normalize` or `-af loudnorm` | | Transcribe | Whisper → text, SRT, VTT | | Separate stems | Demucs (vocals, drums, bass, other) | ## Execution Pattern 1. **Clarify goal** — What format? What loudness? What platform? 2. **Analyze source** — `ffprobe` for codec, sample rate, channels, duration 3. **Process** — FFmpeg/SoX for transformation 4. **Verify** — Check output plays, meets specs, sounds correct 5. **Deliver** — Provide file to user ## Common Requests → Actions | User says | Agent does | |-----------|------------| | "Convert to MP3" | `-acodec libmp3lame -q:a 2` | | "Remove background noise" | Apply highpass/lowpass or dedicated denoiser | | "Normalize for podcast" | `-af loudnorm=I=-16:TP=-1.5:LRA=11` | | "Transcribe this" | Whisper → output SRT/VTT/TXT | | "Extract audio from video" | `-vn -acodec copy` or re-encode | | "Make it smaller" | Lower bitrate: `-b:a 128k` or `-b:a 96k` | | "Speed up 1.5x" | `-af atempo=1.5` | ## Format Quick Reference | Format | Use Case | Quality | |--------|----------|---------| | WAV | Master, editing | Lossless | | FLAC | Archive, audiophile | Lossless compressed | | MP3 | Universal sharing | Lossy, 128-320 kbps | | AAC/M4A | Apple, podcasts | Lossy, efficient | | OGG/Opus | WhatsApp, Discord | Lossy, very efficient | ## Quality Defaults - **Podcast:** -16 LUFS (Spotify), -19 LUFS (Apple) - **Music:** -14 LUFS (Spotify), -16 LUFS (Apple Music) - **MP3 quality:** VBR `-q:a 2` (~190 kbps) or CBR `-b:a 192k` - **Sample rate:** 44.1kHz for music, 48kHz for video sync ## Scope This skill: - Processes audio files user explicitly provides - Runs FFmpeg commands on user request - Does NOT access cloud services without user knowing - Does NOT store files persistently (user manages their files)
GitHub에서 보기