Communitygithub.com

SrigadaAkshayKumar/ai-video-editor

Design the sound effects of an edit in work/audio-plan.json sfx_events — which effect, where, how often — using the measured SFX library (repo assets/sfx, HyperFrames' bundled set, CC0 fetches from Openverse) and the spectral rule that sub-band cues carry structure while vocal-band cues are rationed. Mixed and levelled by tools/mix.mjs. Use when adding or adjusting SFX in any edit.

ai-video-editor 是什麼?

ai-video-editor is a Claude Code agent skill that design the sound effects of an edit in work/audio-plan.json sfx_events — which effect, where, how often — using the measured SFX library (repo assets/sfx, HyperFrames' bundled set, CC0 fetches from Openverse) and the spectral rule that sub-band cues carry structure while vocal-band cues are rationed. Mixed and levelled by tools/mix.mjs. Use when adding or adjusting SFX in any edit.

相容平台~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/SrigadaAkshayKumar/ai-video-editor/tree/HEAD/.claude/skills/sfx

在你喜歡的 AI 中提問

開啟一個已預先載入此 Agent Skill 的新對話。

說明文件

sfx — work/audio-plan.json → sfx_events

{ "sfx_events": [ { "time": 19.8, "sfx": "sub_drop", "reason": "dock-in for the checklist cluster at 20.3" },
                  { "time": 24.2, "sfx": "impact", "gain": 0.8, "reason": "big_statement 'On paper, the jobs exist'" } ] }

Clean-cut times. sfx is a NAME (npm run sfx -- list [<p>] shows them all with length and band). gain can only push a cue further DOWN. Levels are measured, not chosen: every effect lands ~14 dB under this narration with a peak ceiling, a short-transient floor and a length trim (tools/mix.mjs, see work/levels.json). Never "fix" a level by guessing a gain.

The one rule: impact, not irritation

The template measured it: dialogue is ~81% in 300-3400 Hz. A cue that lives ON the voice (whoosh, pop, click, ding) competes with every word however quiet; a cue that lives UNDER it (sub_drop, impact, riser: 74-100% below 120 Hz here) is felt and never masks speech. Its noisy swipe was not fixed by turning it down — it was replaced by a sub-band cue. So:

Sub-band — carries the structure

  • sub_drop — the reframe cue: ~0.5s before the FIRST overlay of a dock/corner/band cluster, and once at the cluster's end. One pair per CLUSTER, never one per chart.
  • impact — lands a punchline: big_statement/word_swap, only the ~6-8 real turning points.
  • riser (1.7s) — builds INTO a chapter_open: start = chapter start − 1.5s. ≤4-5 per video.
  • also: impact-bass-1/2, whoosh-cinematic (HyperFrames set).

Vocal-band — rationed

  • pop, click — an individual reveal inside a cluster, a keyword_chip/annotation. Leave many reveals silent.
  • ding — a checklist "yes" / resolved point, a handful at most.
  • whoosh — movement only (a whip/flash b-roll entry), never within ~6s of another whoosh.
  • typing — only under a term_card or code reveal. clock_tick, glitch, reveal — when the moment literally is that.

Need something else? npm run sfx -- fetch <p> <name> ["literal query"] pulls a CC0 one-shot from Openverse; then npm run sfx -- list <p> to see its band before using it. Keep the whole video to 6-10 distinct effects.

Placement

  • Never two cues within 0.35s; never the same cue within ~4s; no single effect over ~a third of all cues (it becomes a tic — mix.mjs warns).
  • Movement cues follow the picture, accent cues follow meaning. One cue per moment — never a swipe and a whoosh on the same transition.
  • No SFX inside a chapter_open window except the riser into it. No SFX on the payoff line itself — land the line dry, put the cue on the graphic after it.
  • Budget: ~1 cue per 8-10s (talking head), ~1 per 4-6s clustered (faceless: dense in a montage, near-silent under a long calm explanation). If a cue sits on more than ~half the overlays, cut back.

npm run mix -- <p> builds work/mix.wav (finalize re-mixes automatically when the plan changes). You cannot hear it — tell the user to listen for SFX that land wrong or grate.

Individual skills in this repo

This repo contains 9 individual skills — each has its own dedicated page.

SrigadaAkshayKumar/ai-video-editor

Turn a finished AI-edits video (projects/<p>, or several parts of one long video) into YouTube Shorts. Reuses the edit's word timestamps, reads the whole transcript, picks self-contained dialog moments, adds text hooks + comment prompts, renders 9:16 clips with Remotion. Use when the user asks to clip/make shorts from an edited project or a long video.

SrigadaAkshayKumar/ai-video-editor

Choose, fetch and place background music for an edit — Openverse search (commercial-use CC, or CC0-only), music_cues in work/audio-plan.json with fades, offsets and a deliberate silence gap; tools/mix.mjs conditions every bed, sets it to one measured house level under the voice and sidechain-ducks it. Use when adding, swapping or re-levelling music.

SrigadaAkshayKumar/ai-video-editor

Generate and style word-synced captions (Telugu, Hindi, English, code-mixed) as the HyperFrames captions pass over each format's Remotion picture, fix caption text, black captions out under on-screen text, and check caption content at every cut join. Use when creating, restyling, correcting or repositioning captions in any edit.

SrigadaAkshayKumar/ai-video-editor

Editorial rules for turning a raw talking video or voiceover into a clean cut — removing retakes, repetitions, false starts, stutters, fillers, mispronounced words, contradictions/inconsistencies, off-script chatter and dead air — by writing work/edl.json for tools/cut.mjs. Telugu, Hindi, English and code-mixed speech. Use for stage 2 of edit-video or whenever the user asks to tighten/re-cut a video.

SrigadaAkshayKumar/ai-video-editor

Turns a bare voiceover (no presenter on camera) into a finished faceless documentary/explainer for YouTube 16:9 and Instagram 9:16 — clean-cuts the narration, DIRECTS it (story spine, tension curve, hooks, involvement beats), researches the claims and screenshots the real sources, plans a scene track (stock b-roll, stills, news screenshots, pure-graphic beats) with motion graphics, designs music + SFX, renders, captions and credits it. Use when the user gives a voiceover/narration file (or says "faceless video", "make a video from this VO", "documentary").

SrigadaAkshayKumar/ai-video-editor

End-to-end editor for TALKING-HEAD videos (a person on camera) in this repo. Use when the user hands over a raw video with a speaker and wants a finished edit, or asks to continue/resume/redo any stage of an existing talking-head project — clean cut, motion graphics, b-roll, SFX, music, captions, final render, thumbnail. Delivers YouTube 16:9 and Instagram 9:16. Telugu, Hindi, English and code-mixed speech. For a voiceover with no presenter use edit-video-faceless.

SrigadaAkshayKumar/ai-video-editor

Plan the picture of an edit in work/visual-plan.json — motion graphics (28 component types across six stage layouts that dock/shrink/dim the video), b-roll cutaways from Pexels/Pixabay, camera moves, a-roll framing for 9:16 — validate it, QA it with stills in both formats, and render it with Remotion. Use for any titles, charts, stats, lower thirds, diagrams, b-roll or zooms in a talking-head or faceless edit.

SrigadaAkshayKumar/ai-video-editor

Write the upload package for a video this repo edited — YouTube title options, description with chapters, tags, pinned comment, thumbnail text, and the Instagram Reels caption + hashtags for the 9:16 cut — optimised for whichever channel/client the user says it is for (config/brands/ profiles, only when named). Use when the user asks what to title a video, for a description, tags, chapters, "the YouTube metadata", an Instagram caption, or SEO/CTR advice on a finished render.

SrigadaAkshayKumar/ai-video-editor

Design and build a high-CTR thumbnail for a video this repo edited — hook selection, the real face frame pulled from the footage, background cut-out, type set with real fonts (Telugu/Devanagari included), the YouTube 1280x720 thumbnail and the Instagram 1080x1920 cover, an optional image-generation prompt for the background layer, and a feed-size legibility check. Use when the user asks for a thumbnail, thumbnail ideas, a cover image, or how to make the thumbnail more clickable.

相關技能