원클릭으로
speech-to-text
When the user needs to transcribe speech (such as .mp3, .wav, .ogg) into text, use the python_repl tool to generate text.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
메뉴
When the user needs to transcribe speech (such as .mp3, .wav, .ogg) into text, use the python_repl tool to generate text.
Codex 또는 Claude로 설치 이 Prompt를 복사해 Codex, Claude 또는 다른 어시스턴트에 붙여 넣으면 Skill 페이지를 검토하고 설치를 진행할 수 있습니다.
SOC 직업 분류 기준
Schedule reminders and recurring tasks. Actions: add, list, remove or set_context.
Periodic wake-up service that checks HEARTBEAT.md for pending tasks and executes them automatically. Use this skill to write or update HEARTBEAT.md.
Private knowledge base for indexing multimodal files or folders into a knowledge graph, supporting multi-hop graph retrieval
Generate wiki docs + Mermaid diagrams for any codebase. Use when the user asks to document a codebase, generate architecture diagrams, or create a structured wiki for a repo.
Parse an image from a file path to obtain a description, enabling non-multimodal LLMs to have vision capabilities.
When the user needs to transcribe video (such as .mp4, .mkv, .avi) into text, use the python_repl tool to generate text.
| name | speech_to_text |
| description | When the user needs to transcribe speech (such as .mp3, .wav, .ogg) into text, use the python_repl tool to generate text. |
from skills.builtin.core.speech_to_text.scripts import stt
if __name__ == '__main__':
audio_path: str = "{placeholder}" # <- replace with the absolute path of the input audio file
res = stt(audio_path)
print(res)