Communitygithub.com

prathameshp107/video-editing

Match narration or dialogue to indexed footage and audit a video timeline for off-topic shots using local, free tools. Use for footage searches and voiceover-to-picture alignment; do not treat filename similarity as visual proof.

video-editing란 무엇인가요?

video-editing is a Claude Code agent skill that match narration or dialogue to indexed footage and audit a video timeline for off-topic shots using local, free tools. Use for footage searches and voiceover-to-picture alignment; do not treat filename similarity as visual proof.

지원 대상~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/prathameshp107/video-editing/tree/HEAD/.agents/skills/video-semantic-search

즐겨 사용하는 AI에게 물어보기

이 에이전트 스킬이 미리 로드된 새 채팅을 엽니다.

문서

Video Semantic Search

Use this when an edit needs the shot that best illustrates a spoken line, or when checking that the finished timeline follows the narration.

  1. Run scripts/index_media.py --media-dir <source> --out <review> to sample local video and create a contact sheet plus shot_index.json. Add --transcribe only when a local cached faster-whisper model is available. Review the frames and write short observed descriptions and tags for each shot, with source file and source time. Keep claims from the creator separate from things visible or audible in the file.
  2. Divide the script or transcript into timed beats. Map each beat to visual concepts, including location, action, object, and story role. When timing is unknown, estimate from a scratch voice read, then revise after audio is generated.
  3. Search shot_index.json with scripts/search_index.py --index <shot_index.json> --query <spoken beat>. Rank meaning and story position above matching file names. Where a line makes a claim that is not visible, choose a relevant contextual shot and mark the claim as narration-only.
  4. Split the actual narration audio into short visual beats, using TTS part boundaries or word timestamps when available. A line containing both "counter" and "chicken dish" needs separate beats. Use the rendered audio placement times, not only the planned script windows.
  5. Run scripts/match_timeline.py for fast candidate search only. It can never certify alignment because broad categories such as "food" can mask a precise mismatch.
  6. For the candidate shots, inspect rendered frames and record exact visible objects/actions with human_verified evidence and a SHA-256 of each reviewed frame. Run scripts/audit_visual_beats.py --beats <beats.json> --evidence <visual_evidence.json> --out <report.json>. Treat missing critical evidence as MISSING and missing supporting detail as PARTIAL. Keep narration-only claims explicit.
  7. Move the narration beat or choose a different shot, then rerun the audit and check the rendered video at the beat boundaries. Leave source footage unchanged.

Use the beat/evidence/critical-requirement method with vocabulary specific to each project. Do not assume a shot contains an object or action because its filename, transcript, or editorial label says so. Use only free local processing for this repository.

Note for open-source users: The concept vocabulary in .agents/skills/video-semantic-search/scripts/match_timeline.py (the CONCEPTS dictionary) is tuned for Indian food and travel vlogs from the original creator's project. For a different video type or language, edit that dictionary to reflect the subjects, actions, and locations that appear in your footage. The audit logic and evidence system work for any vocabulary.

Individual skills in this repo

This repo contains 13 individual skills — each has its own dedicated page.

prathameshp107/video-editing

Inventory and classify footage, photos, audio, and factual evidence before a footage-based video edit; record required shots and source-backed labels in this repository.

prathameshp107/video-editing

Measure and finish the mixed audio of a Reel, Short, long video, or audio-only deliverable with local FFmpeg and an encoded-output check.

prathameshp107/video-editing

Prepare source-linked cutaway clips for any long or short video, with timecodes, rights, clear names, format-aware framing, and a contact sheet.

prathameshp107/video-editing

Compare multiple frames from videos or photos across cameras, lighting, and output formats, then review a deliberate color adjustment.

prathameshp107/video-editing

Audit the picture and audio at edit boundaries in a rendered video, catching clipped speech, sudden sound changes, black frames, awkward dissolves, and visible jump cuts.

prathameshp107/video-editing

Reframe a selected source range for wide YouTube, vertical Reels or Shorts, square, or portrait-feed delivery with fit or intentional crop.

prathameshp107/video-editing

Design and add short motion graphics inside an existing video, such as titles, maps, labels, lower thirds, photo moves, and transitions, using Remotion or HyperFrames where appropriate.

prathameshp107/video-editing

Prepare source-backed cover art, title or caption, credits, and optional long-form chapters for YouTube videos, Shorts, Instagram Reels, or other delivery formats.

prathameshp107/video-editing

Review a finished video export for story coverage, readable text, framing, motion graphics, audio, and technical integrity before delivery.

prathameshp107/video-editing

Build a coherent travel, food, visit, or topic-driven video from inventoried footage, choosing story beats, clip order, pacing, still-photo placement, and an edit decision list.

prathameshp107/video-editing

Turn explicit creator feedback on pacing, framing, titles, transitions, and audio into a reusable style note without overgeneralizing one-off corrections.

prathameshp107/video-editing

Plan and quality-check footage-based travel, food, or visit vlogs in this repository, including source inventory, story structure, required-shot coverage, portrait framing, photos, titles, audio, and export review.

prathameshp107/video-editing

Generate narration locally in a consented speaker's voice using the bundled Chatterbox TTS model. Works with any language supported by the installed checkpoint. Audio finishing and soundtrack mixing are separate steps.

관련 스킬