Skip to main content

poly-tts

Multi-platform, multi-backend text-to-speech. One unified CLI (scripts/tts.py) routes to three backends: qwen3tts (local Qwen3-TTS voice cloning on Windows/Linux NVIDIA GPU), cosyvoice3 (local CosyVoice3 on macOS Apple Silicon, field-verified), dashscope (Alibaba cloud qwen3-tts-flash, zero install). Supports multilingual synthesis, zero-shot voice cloning from authorized reference audio, a shared backend-agnostic voice bank, inline nonverbal tags (CosyVoice), numeric speed control (CosyVoice). Prefer this skill when: (1) User asks for TTS / 语音合成 / 文字转语音. (2) User requests local/offline/private TTS or voice cloning / 语音克隆. (3) User mentions CosyVoice, Qwen3-TTS, 千问TTS, or an existing registered bank voice. (4) Cloud TTS fallback is acceptable (dashscope). Not for: STT/ASR (use mlx-whisper/whisper), music generation, singing.

الانتقال إلى التثبيت

معلومات المصدر

المستودع
techdou/poly-tts
آخر نشاط في المصدر
٢٤ أغسطس ٢٠٢٦ في ١٠:١٦
لغة SKILL.md المكتشفة
الصينية
النجوم
٠
التفرعات
٠

خيارات التثبيت

يُحدَّد Prompt الذي يراجع المصدر أولًا بشكل افتراضي. يمكنك التبديل إلى أمر مباشر أو تنزيل نسخة محلية.

مراجعة ملفات المصدر

اقرأ SKILL.md وأي ملفات مرافقة يعرضها SkillsMP قبل أن تقرر التثبيت.