人物音色设计(Qwen3-TTS)
SkillMediaWrite character voice parameters for the Qwen3-TTS voice design workflow. Use when designing a dedicated voice for a comic-drama character (voice sample voice_ref); provide a representative line of dialogue (text) and a voice description (voice_description). Outputs text and voice_description.
Available today. Use it from your connected AI after setup.
No other account needed.
Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.
Then ask your AI: use the 人物音色设计(Qwen3-TTS) skill
What this skill tells your AI
The instructions your AI receives, as published by swotatmk/infinite-creation in skills/tts-voice-design/SKILL.md and read by ahel’s review.
为每个出场人物设计一个专属音色样本,供后续视频生成时作为 voice_ref 参考。
输入
- 人物 bible.md 中的 voice 特征(若有)。
- 人物在本章的对白风格(年龄、性别、性格、情绪基调)。
text(朗读文案 · 简短自我介绍)
写 1 句 20~30 字左右的简短自我介绍(约 5 秒,口语化、贴合人物语气),作为人物朗读的音色合成样本。示例:你好呀,我是沈舒妍,喜欢安静的午后和温暖的光,很高兴认识你。
禁止长篇大论:不要塞大段台词、剧情或描写,20~30 字即可;文案过长会生成过长的音色样本。
voice_description(音色描述 · 不限制字数)
描述目标音色,不限制字数,可详细写,覆盖:性别 + 年龄段 + 音色质感(明亮/低沉/沙哑/清亮/温润)+ 语速 + 语气情绪 + 口音/风格。示例:青年男性,嗓音低沉温和,语速偏慢,带一点书卷气的平静口吻。
输出与调用
对每个人物输出 {"text":"...","voice_description":"..."},然后调用 design_voice(asset_id, text, voice_description) 生成音色样本(落盘为该项目 voice_ref)。多个角色连续调用 design_voice,避免中途切回生图工作流造成模型重载。
约束
- 音色与人物年龄/性别/性格一致;不同角色音色尽量区分。
- voice_description 中文、具体,避免「好听」这类空词。
Signals
- GitHub stars
- 28
- Forks
- 3
- Last commit
- Sep 2026
Advanced
- Item type
- skill
- Key
tts-voice-design- Source
- github.com/swotatmk/infinite-creation