Communitygithub.com

prathameshp107/video-editing

Generate narration locally in a consented speaker's voice using the bundled Chatterbox TTS model. Works with any language supported by the installed checkpoint. Audio finishing and soundtrack mixing are separate steps.

¿Qué es video-editing?

video-editing is a Claude Code agent skill that generate narration locally in a consented speaker's voice using the bundled Chatterbox TTS model. Works with any language supported by the installed checkpoint. Audio finishing and soundtrack mixing are separate steps.

Compatible con~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/prathameshp107/video-editing/tree/HEAD/.agents/skills/voiceover-generator

Preguntar en tu IA favorita

Abre un nuevo chat con esta habilidad de agente ya precargada.

Documentación

Voiceover generator

Use this skill when the creator wants spoken narration in a voice they supplied or authorized. The consented reference recording goes under private/voice/; keep the original beside it. The model and local runtime live under private/models/ and private/voice_env/. All private files are excluded from Git. Never publish, commit, or share the reference or model files with a project export.

Write the narration in short, speakable sentences. Check pronunciation of names, places, and technical terms before generating. Generate a new WAV without overwriting previous takes:

# Windows (PowerShell)
.\private\voice_env\Scripts\python.exe scripts/voiceover_generate.py --text-file projects/my-video/voiceover.txt --out projects/my-video/work/voiceover_take_01.wav

# macOS / Linux
python3 scripts/voiceover_generate.py --text-file projects/my-video/voiceover.txt --out projects/my-video/work/voiceover_take_01.wav

Use --dry-run first to verify the reference path, model directory, text split, and output path before loading the model. For a different consented speaker, prepare a clean 3–60 second speech-only recording with scripts/prepare_voice_reference.py, then pass it with --reference. A noisy recording or music under speech can transfer artifacts to the output. Listen to every generated take for missing words, unnatural delivery, and mispronounced names. Regenerate only weak sentences and keep chosen takes in the project folder. Split long scripts into scene-sized files so timing can be adjusted against footage.

The bundled Chatterbox checkpoint is language-specific to whichever model is placed in private/models/. Confirm the model supports your target language and run a short test before a full script. A Hindi checkpoint does not produce correct Marathi or English; install a compatible multilingual checkpoint and verify it. scripts/voiceover_generate.py does not place narration on the video timeline. Use video-semantic-search for picture alignment and audio-finisher after the voice, music, and effects are mixed. All processing stays local and uses free software.

For open-source contributors: the private/ folder is intentionally Git-ignored. To use this skill in a new workspace, obtain a consented reference recording, place a compatible Chatterbox model checkpoint under private/models/, create a virtual environment with the model's dependencies under private/voice_env/, and run --dry-run to confirm the paths.

Individual skills in this repo

This repo contains 13 individual skills — each has its own dedicated page.

prathameshp107/video-editing

Inventory and classify footage, photos, audio, and factual evidence before a footage-based video edit; record required shots and source-backed labels in this repository.

prathameshp107/video-editing

Measure and finish the mixed audio of a Reel, Short, long video, or audio-only deliverable with local FFmpeg and an encoded-output check.

prathameshp107/video-editing

Prepare source-linked cutaway clips for any long or short video, with timecodes, rights, clear names, format-aware framing, and a contact sheet.

prathameshp107/video-editing

Compare multiple frames from videos or photos across cameras, lighting, and output formats, then review a deliberate color adjustment.

prathameshp107/video-editing

Audit the picture and audio at edit boundaries in a rendered video, catching clipped speech, sudden sound changes, black frames, awkward dissolves, and visible jump cuts.

prathameshp107/video-editing

Reframe a selected source range for wide YouTube, vertical Reels or Shorts, square, or portrait-feed delivery with fit or intentional crop.

prathameshp107/video-editing

Design and add short motion graphics inside an existing video, such as titles, maps, labels, lower thirds, photo moves, and transitions, using Remotion or HyperFrames where appropriate.

prathameshp107/video-editing

Prepare source-backed cover art, title or caption, credits, and optional long-form chapters for YouTube videos, Shorts, Instagram Reels, or other delivery formats.

prathameshp107/video-editing

Review a finished video export for story coverage, readable text, framing, motion graphics, audio, and technical integrity before delivery.

prathameshp107/video-editing

Build a coherent travel, food, visit, or topic-driven video from inventoried footage, choosing story beats, clip order, pacing, still-photo placement, and an edit decision list.

prathameshp107/video-editing

Turn explicit creator feedback on pacing, framing, titles, transitions, and audio into a reusable style note without overgeneralizing one-off corrections.

prathameshp107/video-editing

Match narration or dialogue to indexed footage and audit a video timeline for off-topic shots using local, free tools. Use for footage searches and voiceover-to-picture alignment; do not treat filename similarity as visual proof.

prathameshp107/video-editing

Plan and quality-check footage-based travel, food, or visit vlogs in this repository, including source inventory, story structure, required-shot coverage, portrait framing, photos, titles, audio, and export review.

Skills relacionados