乔木音视频 #7 GitHub Stars 312 原文已核对

listenhub-voice

用 ListenHub Voice 从文本或图片生成音频,适合旁白、播客和声音素材制作。

下载视频、音频、字幕和元数据,做可复核素材留档

安装这个 Skill

npx skills add https://github.com/joeseesun/qiaomu-cut-skill --skill listenhub-voice

为什么值得装

推荐它,是因为这是 joeseesun 仓库中可直接追溯到原始 SKILL.md 的条目,文档路径为「vendor/marswaveai-skills/listenhub-voice/SKILL.md」。相比只看仓库名,原文给出了触发场景、执行边界或操作步骤,适合在需要这类能力时直接安装或作为模板改造。

原文里最有价值的点

  • End-to-end audio generation with ListenHub-Voice-1.0 (text / image → audio)。
  • User wants end-到-end 音频 来自 text (结合 sound effects baked in by the model)。
  • User wants a multi-voice dialogue where each line is assigned 到 a different。
  • 文档重点章节:When to Use、When NOT to Use、Purpose。

适合用在

  • 下载、剪辑、转写、配音、生成或整理音视频素材
  • 制作字幕、播客、音乐、短视频和多媒体发布资产