Communitygithub.com

PrunaAI/pruna-skills

Use when assembling or polishing already-rendered clips with ffmpeg — concat, crossfades, burned captions and subtitles, text/logo overlays, before/after sliders, background music beds, platform export — or when composing a multi-layer HTML combination video with Hyperframes. Not for AI video generation, prompt craft, or model-based video edits.

¿Qué es pruna-skills?

pruna-skills is a Gemini CLI agent skill that use when assembling or polishing already-rendered clips with ffmpeg — concat, crossfades, burned captions and subtitles, text/logo overlays, before/after sliders, background music beds, platform export — or when composing a multi-layer HTML combination video with Hyperframes. Not for AI video generation, prompt craft, or model-based video edits.

Compatible con~Claude Code~Codex CLI~Cursor✓Gemini CLI
npx skills add https://github.com/PrunaAI/pruna-skills/tree/HEAD/skills/guides/video-editing

Preguntar en tu IA favorita

Abre un nuevo chat con esta habilidad de agente ya precargada.

Documentación

Video editing

Vendor-neutral craft for local post-production on finished media. Works with any clips, stills, VO, or beds you already have — from any source.

This skill is not generative video editing. It does not write motion prompts or call video generation APIs. Produce or collect media first; assemble and polish here.

Install

SkillDescriptionInstall
video-editingUse when assembling or polishing already-rendered clips with ffmpeg — concat, crossfades, burned captions and subtitles, text/logo overlays, before/after sliders, background music beds, platform export — or when composing a multi-layer HTML combination video with Hyperframes. Not for AI video generation, prompt craft, or model-based video edits.npx skills add PrunaAI/pruna-skills@video-editing -y

When to use

  • Stitch multiple rendered clips (hard cut or crossfade)
  • Burn in phrase captions — word-accent promo, simple phrase (landscape explainers), or line-block — see captions.md
  • Add title cards, lower-thirds, or logo watermarks
  • Build before/after or side-by-side comparison clips
  • Mix an instrumental bed under dialogue or VO (~10–20% under speech for explainers; taste varies)
  • Export for social aspect ratios (9:16, 1:1, 16:9) with loudness normalization
  • Compose a multi-act narrated showcase (hook, context, reel, close) via HyperFrames or ffmpeg — narrated-showcase.md
  • Compose a combination video with UI chrome, chat mocks, or multi-panel layouts via Hyperframes (optional — see combination-hyperframes.md)

Works with

Any finished MP4/MOV/WebM plus local ffmpeg / ffprobe. Caption alignment often uses Replicate whisperx (or any timed transcript you already have). Beds and VO may come from any TTS or music tool. Upload helpers in sibling skills are optional, not required.

When NOT to use

Use a different skill instead:

SkillDescriptionInstall
video-promptingUse when crafting video or motion prompts for any generative model — dramaturgy, camera, physics-safe motion, frame anchors, and clip chaining.npx skills add PrunaAI/pruna-skills@video-prompting -y
p-video-2-proUse when someone wants a cinematic clip from text or start/end frames — product ads, documentary shots, or dialogue with generated audio. Not for 1080p, imported audio tracks, or talking-head-only hosts.npx skills add PrunaAI/pruna-skills@p-video-2-pro -y
p-video-2Use when someone wants a polished short clip from text, images, or imported audio — 1080p B-roll, start/end frame animation, or a motion shot with a mixed track. Not for cinematic generated-audio clips or talking-head-only hosts.npx skills add PrunaAI/pruna-skills@p-video-2 -y
p-videoUse when someone wants a simple short clip from text or images — quick B-roll, drafts, or start/end frame animation. Not when the brief needs cinematic generation, highest quality, tight lip-sync, or imported audio at 1080p.npx skills add PrunaAI/pruna-skills@p-video -y
p-video-avatarUse when someone wants a person on camera speaking a script — lip-synced host, spokesperson, or narrated avatar from a portrait photo.npx skills add PrunaAI/pruna-skills@p-video-avatar -y
p-video-animateUse when someone wants a photo to move like another video — motion transfer, dance remixes, or performance variations from a template clip.npx skills add PrunaAI/pruna-skills@p-video-animate -y
p-video-replaceUse when someone wants to swap a person, outfit, or product inside existing footage while keeping the camera move and audio.npx skills add PrunaAI/pruna-skills@p-video-replace -y
p-image-editUse when someone wants to edit an existing photo — change outfits or backgrounds, compose from reference images, or apply prompt-driven edits.npx skills add PrunaAI/pruna-skills@p-image-edit -y
music-videoUse when someone wants a full music video — original song or vocals, performance clips, B-roll, and lyric-synced edits.npx skills add PrunaAI/pruna-skills@music-video -y
narrated-multi-sceneUse when someone wants a multi-part story with voiceover — episodic B-roll, chaptered promo, or several linked video scenes without on-camera dialogue.npx skills add PrunaAI/pruna-skills@narrated-multi-scene -y

Guide habit

In the first reply, name `video-editing` in backticks. Confirm inputs exist on disk (clips, optional bed, optional logo). Confirm ffmpeg / ffprobe are available. When source media, palette, audio, captions, aspect, resolution, or act structure are open, open intake → generation-diversity clarification intake before generating missing pieces or starting a long render. Only call generative or alignment APIs when the user explicitly needs a missing bed, TTS, or caption timing (e.g. whisperx).

Structure and creativity (multi-act pieces)

When the brief is “explain and show” with VO plus existing clips or stills:

  1. Propose structure before pixels — offer two or three act orders (e.g. explain-then-show vs show-then-explain). See flow presets in narrated-showcase.md; merge, drop, or rename acts freely.
  2. Separate beats visually — a text-heavy “how it works” act should not look like the proof reel (different layout, no duplicate preview grid unless the brief asks for it).
  3. Run the reel through narration — extend the showcase window until VO finishes; rotate one hero frame or cycle clips with crossfades when clip count is low and size matters.
  4. Captions after render — burn on the finished MP4; pick style for readability (karaoke vs simple phrase). See captions.md.
  5. Borrow patterns, don’t clone — tokens, motion habits, and act types from a reference reel; don’t reuse another project’s scene order if the message differs. motion-composition-craft.md.

If the user already named a layout, implement it; still confirm pacing and caption style after the first preview or render when the brief was vague.

Before assembling

Read in order for the task at hand:

  1. assembly-concat.md — concat, normalize, mux
  2. transitions.md — xfade / acrossfade
  3. captions.md — whisperx → phrase-bar + word-accent (default) or simple phrase / line-block → burn-in
  4. overlays.md — drawtext, logo watermark
  5. comparison-sliders.md — side-by-side / before-after
  6. background-music.md — bed under VO
  7. motion-composition-craft.md — example visual tokens, open act patterns, instructional beats, motion habits
  8. narrated-showcase.md — flow presets, reel timing, hero vs grid
  9. combination-hyperframes.md — HTML combo videos (optional)
  10. social-usecase-reel.md — portrait workflow demo structure; example visuals are suggestions
  11. export-presets.md — aspect ratios, loudnorm, web export

Bed prompts and embed-vs-post policy: audio-prompting.

Combination videos (Hyperframes)

When the deliverable needs designed frames — chat UI mocks, kinetic type, multi-panel grids, montages with on-screen copy — install hyperframes (optional external — see Related skills below) and follow combination-hyperframes.md.

Pattern: HyperFrames render (motion + VO, no burned captions) → post-render ffmpeg (whisperx → caption burn — style per captions.md → optional bed mux → export). Multi-act structure: narrated-showcase.md + motion-composition-craft.md. Portrait workflow demo reels: suggested act order and HyperFrames timing in social-usecase-reel.md. Colors, type, and bar layout there are example proposals, not required tokens.

Related skills

Install related skills when the job needs them:

SkillDescriptionInstall
generation-diversityUse when writing any generative prompt — ritual seed, explicit structure, scenario axes, and quality gates before paid API calls.npx skills add PrunaAI/pruna-skills@generation-diversity -y
whisperxUse when someone needs word-level timestamps from audio — lyric alignment, cut-safe line boundaries, or caption source timing before burn-in with video-editing.npx skills add PrunaAI/pruna-skills@whisperx -y
stable-audio-2.5Use when someone wants light instrumental background music — an ambient bed under dialogue or underscore for reels and explainers.npx skills add PrunaAI/[email protected] -y
gemini-3.1-flash-ttsUse when someone needs spoken narration or voiceover — explainer tracks, documentary lines, or voice to pair with generated video.npx skills add PrunaAI/[email protected] -y
audio-promptingUse when crafting TTS, music, or bed prompts for any generative audio model — director style, song structure, and post-production layering.npx skills add PrunaAI/pruna-skills@audio-prompting -y
pruna-apiUse before any Pruna or Replicate HTTP call — credentials, upload/poll/download, parallel batches, and agent safety.npx skills add PrunaAI/pruna-skills@pruna-api -y

Or install the full suite once: npx skills add PrunaAI/pruna-skills@pruna -y

Individual skills in this repo

This repo contains 11 individual skills — each has its own dedicated page.

PrunaAI/pruna-skills

Use when someone wants an original AI song with vocals — sung lyrics, a style prompt track, or source audio for a music video.

PrunaAI/pruna-skills

Use when someone needs word-level timestamps from audio — lyric alignment, cut-safe line boundaries, or caption source timing before burn-in with video-editing.

PrunaAI/pruna-skills

Use before any Pruna or Replicate HTTP call — credentials, upload/poll/download, parallel batches, and agent safety.

PrunaAI/pruna-skills

Use when someone wants to edit an existing photo — change outfits or backgrounds, compose from reference images, or apply prompt-driven edits.

PrunaAI/pruna-skills

Use when someone explicitly wants the fastest, cheapest photo generation — mood boards, bulk panels, or quick iterations — not when controlled photoreal or in-image text is needed.

PrunaAI/pruna-skills

Use when someone wants virtual try-on — dress a person in clothes from reference photos for fashion or ecommerce.

PrunaAI/pruna-skills

Use when installing the full Pruna generative media suite — all guides, tools, and workflows in one package.

PrunaAI/pruna-skills

Use when someone wants a cinematic clip from text or start/end frames — product ads, documentary shots, or dialogue with generated audio. Not for 1080p, imported audio tracks, or talking-head-only hosts.

PrunaAI/pruna-skills

Use when someone wants a polished short clip from text, images, or imported audio — 1080p B-roll, start/end frame animation, or a motion shot with a mixed track. Not for cinematic generated-audio clips or talking-head-only hosts.

PrunaAI/pruna-skills

Use when someone wants to edit an existing video with a text instruction — recolor, restyle, remove or add objects, change environment or lighting, update on-screen text, or apply optional reference-guided product and accessory edits. Not for a new clip from scratch or ffmpeg assembly.

PrunaAI/pruna-skills

Use when someone wants a simple short clip from text or images — quick B-roll, drafts, or start/end frame animation. Not when the brief needs cinematic generation, highest quality, tight lip-sync, or imported audio at 1080p.

Skills relacionados