乔木仓库

listenhub-voice

用 ListenHub Voice 从文本或图片生成音频,适合旁白、播客和声音素材制作。

分类
乔木音视频
榜单排名
#7
GitHub Stars
312

它能做什么

中文摘要

「listenhub-voice」来自 joeseesun/qiaomu-cut-skill 的公开 SKILL.md。原文定位是:用 ListenHub Voice 从文本或图片生成音频,适合旁白、播客和声音素材制作。它适合下载视频、音频、字幕和元数据,做可复核素材留档。

为什么推荐

推荐理由

推荐它,是因为这是 joeseesun 仓库中可直接追溯到原始 SKILL.md 的条目,文档路径为「vendor/marswaveai-skills/listenhub-voice/SKILL.md」。相比只看仓库名,原文给出了触发场景、执行边界或操作步骤,适合在需要这类能力时直接安装或作为模板改造。

什么时候用

适用场景

  • 下载视频、音频、字幕和元数据,做可复核素材留档
  • 下载、剪辑、转写、配音、生成或整理音视频素材
  • 制作字幕、播客、音乐、短视频和多媒体发布资产

使用前先看

主要亮点

  • 01

    End-to-end audio generation with ListenHub-Voice-1.0 (text / image → audio)。

  • 02

    User wants end-到-end 音频 来自 text (结合 sound effects baked in by the model)。

  • 03

    User wants a multi-voice dialogue where each line is assigned 到 a different。

  • 04

    文档重点章节:When to Use、When NOT to Use、Purpose。

原始文档

原文摘录

End-to-end audio generation with ListenHub-Voice-1.0 (text / image → audio).