RabiSpeech IndexTTS2

SkillMedia

Lets your agent generate Chinese speech and cloned voices locally using the IndexTTS2 text-to-speech model.

Available today. Use it from your connected AI after setup.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

Then ask your AI: use the RabiSpeech IndexTTS2 skill

About this skill

Generates Chinese speech or voice-cloned audio via the local IndexTTS2 model through RabiSpeech. Use when the user specifies IndexTTS2, a Chinese reference audio, emotion/style instructions, a persona voice, or local WAV verification. Do not directly run the skill's worker, start OumuQ, or call clou

What this skill tells your AI

The instructions your AI receives, as published by vb2250158/rabiroute in skills/indextts2-audio/SKILL.md and read by ahel’s review.

IndexTTS2 由 RabiSpeech 管理,模型 id 为 local-tts/indextts2。人格参考音频位于 RabiRoute/data/roles/<RoleId>/voice/,生成结果不得写回参考库。

流程

先读 TTS 路由。speechBaseUrl 必须从当前服务配置取得;下例中的 $speechBaseUrl 由该发现步骤赋值。

  1. 检查 GET <speechBaseUrl>/v1/models/local-tts/indextts2。
  2. 传入 voice=<RoleId>;RabiSpeech 从人格 voice index 选择并缓存参考音频。
  3. 需要情绪或语气时用 instructions 描述,保持简短、可执行。
  4. 验证用 play=false 并保存返回 WAV;对话用 play=true、session_id,进入全局 FIFO。
$body = @{
  model = 'local-tts/indextts2'
  input = '你好,这是 IndexTTS2 本地语音。'
  voice = '<RoleId>'
  response_format = 'wav'
  speed = 1.0
  language = 'zh'
  instructions = '自然、温和、清楚。'
  play = $false
} | ConvertTo-Json -Compress
Invoke-WebRequest -Method Post -Uri ($speechBaseUrl.TrimEnd('/') + '/v1/audio/speech') -ContentType 'application/json' -Body $body -OutFile '.\indextts2.wav'

不要从技能目录运行旧 indextts2_worker.py;模型环境、CUDA、参考音频组合和输出缓存统一属于 RabiSpeech。

Signals

GitHub stars
502
Last commit
Sep 2026
Advanced
Item type
skill
Key
indextts2-audio
Source
github.com/vb2250158/rabiroute