Audio 模型创建

其他 社区
解读按原文结构重写,命令、链接、术语均保留;右侧可核对作者原始 SKILL.md

方法与流程

  • Workflow:原文以此为独立章节。

适用与边界

  • When to Use Which Tool:tts — When the user wants to generate speech from text using a named speaker (Vivian, Ryan, etc.). Supports English and Chinese. voiceclone — When the user wants to clone a specific voice from a reference audio file and generate new speech in that voice. If…
  • 当前原文没有单列不适用场景或限制。

原文中的明确线索

  • 要点:「tts」、「voiceclone」、「mono 24kHz 16-bit WAV」、「Generate speech audio from text, or clone a voice from a reference audio file.」、「{baseDir}/scripts/tts — Text-to-speech generation with named speakers.」、「{baseDir}/scripts/models/Qwen3-TTS-12Hz-0.6B-CustomVoice — Named speaker TTS (0.6B parameters).」、「Pre-packaged reference audio files for voice cloning are available at {baseDir}/scripts/referenceaudio/.」、「Available reference speakers: trump, elonmusk.」
  • 文件与命令{baseDir}/scripts/tts{baseDir}/scripts/voiceclone{baseDir}/scripts/models/Qwen3-TTS-12Hz-0.6B-CustomVoice{baseDir}/scripts/models/Qwen3-TTS-12Hz-0.6B-Base{baseDir}/scripts/referenceaudio/{baseDir}/scripts/referenceaudio/<speakername>.wav{baseDir}/scripts/referenceaudio/<speakername>.txttrump

流狐整理:以上内容来自当前 SKILL.md 的章节与原词;未补写作者没有声明的工具、兼容性或能力。

流狐档案 作者与许可取自来源;运行、权限和网络为流狐检测或估算
流狐分类
通用
作者声明 Agent
未找到明确声明;不据此推断已兼容或已测试
静态检查
88 / 100 · 启发式扫描,不代表运行安全
作者 / 版本 / 许可
@second-state · 未声明 license
流狐 Token 估算
低消耗
流狐接入估算
即装即用
是否需要外部 API Key
未发现要求
检测到的系统要求
macOS · Linux
底层运行要求
未声明
检测到的文件与系统行为
  • 只读
  • 允许写入 / 修改
检测到的网络行为
仅限本地
安装命令数
无(仅作为资料)

档案由构建时根据 SKILL.md 与安装命令自动衍生,可能与作者实际意图存在差异。

需要注意: 未限定 allowed-tools,默认拥有全部工具权限。

输出预览 audio-tts.preview
# 4. Return the Output

- `tts` produces `output.wav`
- `voice_clone` produces `output_voice_clone.wav`

讨论

基于 GitHub Discussions。登录 GitHub 即可参与讨论、点赞、订阅更新。