视频配音生成
把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 外发说明:每段旁白文字会发给所选 TTS 服务(MiMo / Fish Audio / 用户自托管的 IndexTTS),--voice-ref 的参考音频会发给 MiMo。 另含实验性的英译中 dub 路径:只在显式选择 dub 模式并传
as observed 2026-10-04T09:46:17.065Z- Identifier
video-voiceover- Source
- ClawHub
- Version observed
- 1.0.8
- Source repository
- not published
- Repository observation
- No source repository listed
- First observed here
- 2026-10-03T08:46:10.142Z
- Observations recorded
- 2
- Installs (reported upstream)
- 0
- Weekly downloads (upstream)
- 88
- Declared license
- MIT-0
Observation history
2026-10-04T09:46:17.065Z
Fields that differed: changelog latestVersion summary
| Field | Before | After |
|---|---|---|
changelog |
"- Improved strict mode: old autoshrink caches are no longer used when strict approval-protection is enabled; only exact matches with the same strategy reuse cache.\n- Default over | "**Summary: Expanded privacy and data externalization notes; added experimental dub mode documentation.**\n\n- SKILL.md now clearly details all external network requests and exactl |
latestVersion |
"1.0.1" | "1.0.8" |
summary |
"把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 | "把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 |
Correction
If you maintain this extension and believe anything above is inaccurate, request a correction. Corrections are published, and disputed entries are marked as disputed while under review.