它能做什么
中文摘要
「listenhub-voice」来自 joeseesun/qiaomu-cut-skill 的公开 SKILL.md。原文定位是:用 ListenHub Voice 从文本或图片生成音频,适合旁白、播客和声音素材制作。它适合下载视频、音频、字幕和元数据,做可复核素材留档。
为什么推荐
推荐理由
推荐它,是因为这是 joeseesun 仓库中可直接追溯到原始 SKILL.md 的条目,文档路径为「vendor/marswaveai-skills/listenhub-voice/SKILL.md」。相比只看仓库名,原文给出了触发场景、执行边界或操作步骤,适合在需要这类能力时直接安装或作为模板改造。
什么时候用
适用场景
- 下载视频、音频、字幕和元数据,做可复核素材留档
- 下载、剪辑、转写、配音、生成或整理音视频素材
- 制作字幕、播客、音乐、短视频和多媒体发布资产
使用前先看
主要亮点
- 01
End-to-end audio generation with ListenHub-Voice-1.0 (text / image → audio)。
- 02
User wants end-到-end 音频 来自 text (结合 sound effects baked in by the model)。
- 03
User wants a multi-voice dialogue where each line is assigned 到 a different。
- 04
文档重点章节:When to Use、When NOT to Use、Purpose。
原始文档
原文摘录
End-to-end audio generation with ListenHub-Voice-1.0 (text / image → audio).