人物音色设计(Qwen3-TTS)

SkillMedia

Write character voice parameters for the Qwen3-TTS voice design workflow. Use when designing a dedicated voice for a comic-drama character (voice sample voice_ref); provide a representative line of dialogue (text) and a voice description (voice_description). Outputs text and voice_description.

Available today. Use it from your connected AI after setup.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

Then ask your AI: use the 人物音色设计(Qwen3-TTS) skill

What this skill tells your AI

The instructions your AI receives, as published by swotatmk/infinite-creation in skills/tts-voice-design/SKILL.md and read by ahel’s review.

为每个出场人物设计一个专属音色样本,供后续视频生成时作为 voice_ref 参考。

输入

  • 人物 bible.md 中的 voice 特征(若有)。
  • 人物在本章的对白风格(年龄、性别、性格、情绪基调)。

text(朗读文案 · 简短自我介绍)

写 1 句 20~30 字左右的简短自我介绍(约 5 秒,口语化、贴合人物语气),作为人物朗读的音色合成样本。示例:你好呀,我是沈舒妍,喜欢安静的午后和温暖的光,很高兴认识你。

禁止长篇大论:不要塞大段台词、剧情或描写,20~30 字即可;文案过长会生成过长的音色样本。

voice_description(音色描述 · 不限制字数)

描述目标音色,不限制字数,可详细写,覆盖:性别 + 年龄段 + 音色质感(明亮/低沉/沙哑/清亮/温润)+ 语速 + 语气情绪 + 口音/风格。示例:青年男性,嗓音低沉温和,语速偏慢,带一点书卷气的平静口吻。

输出与调用

对每个人物输出 {"text":"...","voice_description":"..."},然后调用 design_voice(asset_id, text, voice_description) 生成音色样本(落盘为该项目 voice_ref)。多个角色连续调用 design_voice,避免中途切回生图工作流造成模型重载。

约束

  • 音色与人物年龄/性别/性格一致;不同角色音色尽量区分。
  • voice_description 中文、具体,避免「好听」这类空词。

Signals

GitHub stars
28
Forks
3
Last commit
Sep 2026
Advanced
Item type
skill
Key
tts-voice-design
Source
github.com/swotatmk/infinite-creation