Tts skill
Multi-engine text-to-speech skill. Supports Qwen3-TTS local voice cloning, VoiceCraft online TTS, and OpenAI TTS.From its SKILL.md
npx -y skills add aiskillstore/marketplace --skill tts-skillAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
One thing to look at
- no licenseNo license file was found in the repository. Code published without one is not open source by default, so using it at work is a question for whoever answers licensing questions where you are.
SKILL.md
2.9 KB, 802 tokens by cl100k_base, as published. Nobody here has run it
ποΈ TTS-Skill β Multi-Engine Text-to-Speech
TTS-Skill provides a single entrypoint for generating speech using multiple backends, with consistent output naming and progress feedback for long-running jobs.
Engines
- qwen3-tts: local voice cloning with a reference audio + transcript
- edge-tts: online voices with speed/pitch/style controls
- openai-tts: OpenAI speech generation via API
Command Syntax
/tts-skill [engine] [text] --voice [voice-keyword] [other options]
If you use the Python entrypoint:
python tts-skill.py [engine] [text] --voice [voice-keyword]
Text Input
Pass text as a positional argument, or use --text-file / -f to read from a file.
Example:
python tts-skill.py qwen3-tts --text-file "input\\text.txt" --voice ε―ε°ε°ζ
Notes:
--text-filesupports relative and absolute paths; relative paths are resolved from your current working directory- If both positional text and
--text-fileare provided,--text-filetakes priority - UTF-8 is recommended (UTF-8 BOM is supported); on decode error it falls back to GBK
You can also call engine scripts directly:
python engines/qwen3-tts-cli.py --text-file "input\\text.txt" --voice ε―ε°ε°ζ
python engines/edge-tts-cli.py --text-file "input\\text.txt" --voice xiaoxiao
python engines/openai-tts-cli.py --text-file "input\\text.txt" --voice alloy
Local Voice Assets (Qwen3-TTS)
To add a clone voice, put a matching pair of files in assets/:
assets/Lei.wav
assets/Lei.txt
Supported audio formats: .wav, .mp3, .m4a, .flac.
Then:
python tts-skill.py qwen3-tts "ζ΅θ―ζζ¬" --voice Lei
Output
If --output is not provided:
- Output directory:
output/ - Filename pattern:
YYYYMMDD_HHMMSS_<first-6-chars>.<ext>
Progress & Timing (Qwen3-TTS)
Qwen3-TTS jobs print a live progress bar with ETA. After completion, tts-skill.py prints:
- total runtime
- total chars and Chinese chars
- average seconds per Chinese character (or per char if no Chinese)
Project Layout
tts-skill/
βββ .trae/
β βββ plans/
βββ assets/
β βββ Lei.txt
β βββ ε―ε°ε°ζ.txt
β βββ εΈιθ¨.txt
β βββ θ΅΅δΏ‘.txt
βββ engines/
β βββ edge-tts-cli.py
β βββ edge-tts.config
β βββ openai-tts-cli.py
β βββ openai-tts.config
β βββ qwen3-tts-cli.py
β βββ qwen3-tts.config
βββ input/
β βββ text.txt
βββ output/
βββ tts-skill.py
βββ INSTALL.md
βββ INSTALL.zh-CN.md
βββ README.md
βββ README.zh-CN.md
βββ SKILL.md
βββ SKILL.zh-CN.md
Chinese Spec
See SKILL.zh-CN.md.
What ships with it: 19 files
206.8 KB alongside SKILL.md, 4 of them executable
assets/
- Lei.txt41 B
- ε―ε°ε°ζ.txt605 B
- εΈιθ¨.txt344 B
- θ΅΅δΏ‘.txt277 B
docs/
- CHANGELOG.md645 B
- INSTALL.zh-CN.md7.3 KB
- README.zh-CN.md7.1 KB
- SKILL.zh-CN.md6.7 KB
engines/
- edge-tts-cli.pyruns10.5 KB
- edge-tts.config1.3 KB
- openai-tts-cli.pyruns9.7 KB
- openai-tts.config1.7 KB
- qwen3-tts-cli.pyruns18.3 KB
- qwen3-tts.config1.9 KB
input/
- text.txt27 B
- INSTALL.md2.7 KB
- README.md2.0 KB
- skill-report.json123.7 KB
- tts-skill.pyruns11.9 KB