AgentStack
SKILL verified MIT Self-run

Video Voiceover

skill-worldwonderer-video-recap-skills-video-voiceover · by worldwonderer

>

No reviews yet
0 installs
11 views
0.0% view→install

Install

$ agentstack add skill-worldwonderer-video-recap-skills-video-voiceover

✓ scanned · ✓ verified — works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

Are you the author of Video Voiceover? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

What this does

Reads a timestamped narration script and synthesizes one audio clip per segment, fitting speech to each segment's time slot (dynamic rate), then records placement metadata. The only engine is MiMo TTS (mimo-v2.5-tts).

Requirements

export MIMO_API_KEY=***         # MiMo TTS (or a TTS-specific MIMO_TTS_API_KEY)

Input contract

work_dir/narration.json — segments with start / end / narration (+ optional pause_after_ms, overlaps_speech). Times are the output-timeline seconds the audio will be placed at. In the orchestrated cut-mode flow, the agent writes narration.json directly against the output timeline, and the orchestrator passes it here. In the legacy direct-cut path, narration_mapped.json may be passed explicitly instead.

> Running the scripts below — the scripts/… paths are relative to this skill's own directory (the folder containing this SKILL.md). Claude Code runs commands from there, so they work as written. If your harness runs commands from the project root instead (opencode / Codex / OpenClaw commonly do), prefix this skill's absolute directory — e.g. /scripts/…, using the directory your harness reports when it loads the skill. The scripts self-locate from their own path, so once started by the correct path they resolve their sibling skills and assets regardless of the working directory.

Run

python3 scripts/voiceover.py --work-dir  --narration  [--mimo-voice 冰糖]

For direct one-off use, omitting --narration reads work_dir/narration.json. Pass --narration work_dir/narration_mapped.json explicitly only for the legacy direct-cut path; the video-recap orchestrator always passes narration.json.

Output contract

  • tts_segments/*.wav — one synthesized clip per narration segment.
  • tts_meta.json{segments: [...], engine, narration} where each segment carries its

audio_path, timing, pause_after_ms, and placement fields consumed by video-assemble. When --allow-partial-tts lets a run continue past failed segments it also carries partial: true and failures: [{index,start,end,text,error}] so missing lines stay visible (a clean run carries partial: false and failures: []).

Notes

  • Re-runs safely reuse only matching per-segment audio; edited narration or TTS settings regenerate the affected WAVs.
  • TTS_WORKERS, TTS_TIMEOUT, TTS_RETRIES, ALLOW_PARTIAL_TTS tune throughput/robustness.
  • Dub mode has its own deterministic gate: dub_lint.json blocks empty/overlapping/out-of-range

translation lines BEFORE voiceclone spend, and dub_review.json scaffolds fidelity/tone/timing/ platform-fit review. dub.py --stage lint|review and dub.py --print-schema expose them directly.

What this skill does NOT do

  • Does NOT write or edit narration text.
  • Does NOT mux, duck, or render subtitles — that is video-assemble.
  • Does NOT analyze the video or choose timestamps — it voices the segments it is given.

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet — be the first.

Versions

  • v0.1.0 Imported from the upstream source.