TheBotique

视频配音生成

把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 外发说明:每段旁白文字会发给所选 TTS 服务(MiMo / Fish Audio / 用户自托管的 IndexTTS),--voice-ref 的参考音频会发给 MiMo。 另含实验性的英译中 dub 路径:只在显式选择 dub 模式并传

as observed 2026-10-04T09:46:17.065Z
Identifier
video-voiceover
Source
ClawHub
Version observed
1.0.8
Source repository
not published
Repository observation
No source repository listed
First observed here
2026-10-03T08:46:10.142Z
Observations recorded
2
Installs (reported upstream)
0
Weekly downloads (upstream)
88
Declared license
MIT-0

Observation history

2026-10-04T09:46:17.065Z

Fields that differed: changelog latestVersion summary

FieldBeforeAfter
changelog "- Improved strict mode: old autoshrink caches are no longer used when strict approval-protection is enabled; only exact matches with the same strategy reuse cache.\n- Default over "**Summary: Expanded privacy and data externalization notes; added experimental dub mode documentation.**\n\n- SKILL.md now clearly details all external network requests and exactl
latestVersion "1.0.1" "1.0.8"
summary "把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 "把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。

Correction

If you maintain this extension and believe anything above is inaccurate, request a correction. Corrections are published, and disputed entries are marked as disputed while under review.