Skip to main content

speak

Provider-agnostic text-to-speech output with audio caching. Converts text to audio and plays it immediately. Caches generated audio to avoid redundant API calls on repeated phrases. Falls back through available TTS providers (elevenlabs, macOS say). Use as a composable voice output primitive from any skill, agent, or workflow. TRIGGER when: user says "say out loud", "say aloud", "announce", "read out", "tell me out loud", or any phrasing that implies audible/voice output rather than text. Also trigger when user says "say X" at the end of a task request (e.g., "do X and when done say Done") — this means spoken output, not typed. DO NOT TRIGGER when: "say" is used figuratively ("let's say we have...") or means "write/type" in context.

Jump to install

Source facts

Repository
orakitine/toolbox
Last source activity
June 22, 2026 at 19:27
Detected SKILL.md language
English
Stars
0
Forks
0

Install options

The review-first prompt is selected by default. You can switch to a direct command or download a local copy.

Review the source files

Read SKILL.md and any companion files shown by SkillsMP before deciding whether to install.