Communitygithub.com

ido6/ai-video-skills

Design cinematic image and video prompts for genmedia. Use this for shot language, camera movement, lighting, lens choices, color grade, film texture, scene blocking, and production-ready visual direction.

ai-video-skills 是什麼?

ai-video-skills is a Codex agent skill that design cinematic image and video prompts for genmedia. Use this for shot language, camera movement, lighting, lens choices, color grade, film texture, scene blocking, and production-ready visual direction.

相容平台~Claude Code✓Codex CLI~Cursor
npx skills add https://github.com/ido6/ai-video-skills/tree/HEAD/skills/cinematography

在你喜歡的 AI 中提問

開啟一個已預先載入此 Agent Skill 的新對話。

說明文件

Local compatibility and verification (2026-10-04)

User instructions and selected provider/model take precedence over the examples and defaults below. Provider names alone do not identify a model version. Treat endpoint IDs, prices, durations, reference limits, tool availability, and quality rankings in this document as reference examples; verify them against the currently connected tool schema or official provider documentation before execution. Never silently substitute a provider or model.

Installing or updating this skill does not authorize generation, uploads, purchases, or retries. Follow the user's generation approval and budget instructions. Keep existing idoGen mandatory image/video/acting skill requirements. Inspect actual output for identity, spatial continuity, physical motion, and requested format before calling a generation successful. Documentation and installation checks are not output-quality tests.

Cinematography with genmedia

Use this skill when the user needs cinematic direction, not generic "make it cinematic" prompting. Load references as needed:

  • references/shot-language.md
  • references/lighting-lens-color.md
  • references/examples.md
  • references/camera-directing-playbook.md: Ido's own playbook. Load it for any 9:16 ad, any previs/blockout-driven shot plan, handheld-with-tension direction, lens/FOV choice on a rotated full-frame gate, real camera packages, safe zones, or packshot rules. Where it conflicts with the generic references above, it wins.

Load model-routing alongside this skill for default endpoint choices.

Write concrete visual direction. Avoid empty prestige words and em dashes.

Inputs to collect

Ask only for what affects the shot:

  • Subject and action.
  • Medium: still image, video, image-to-video, edit, storyboard frame.
  • Genre and mood.
  • Framing: close-up, medium, wide, overhead, POV, profile, locked-off.
  • Camera motion for video: push-in, dolly, tracking, handheld, crane, drone.
  • Lens feel: wide, normal, telephoto, macro, shallow or deep focus.
  • Lighting: natural, practical, studio, noir, high key, low key, backlit.
  • Output: aspect ratio, duration, first frame, last frame, download path.
  • Preferred model, if the user wants a specific cinematography model or quality/cost profile.

Genmedia workflow

  1. Start from routed endpoint IDs.

    genmedia models --endpoint_id openai/gpt-image-2 --json
    genmedia models --endpoint_id fal-ai/nano-banana-pro --json
    genmedia models --endpoint_id bytedance/seedance-2.0/text-to-video --json
    genmedia models --endpoint_id bytedance/seedance-2.0/image-to-video --json
    genmedia models --endpoint_id xai/grok-imagine-video/text-to-video --json
    

    Use text search only as fallback discovery for a missing camera-control role:

    genmedia models "cinematic video generation camera movement" --json
    genmedia docs "video generation camera movement prompt" --json
    
  2. Inspect schema and use only supported controls.

    genmedia schema <endpoint_id> --json
    genmedia pricing <endpoint_id> --json
    
  3. Upload references when using image-to-video, first frame, last frame, style reference, or character/product continuity.

    genmedia upload ./frame.png --json
    
  4. Run stills with direct download.

    genmedia run <endpoint_id> \
      --prompt "<cinematography prompt>" \
      --download "./outputs/cinema/{request_id}_{index}.{ext}" \
      --json
    
  5. Run video async.

    genmedia run <endpoint_id> \
      --prompt "<shot prompt>" \
      --image_url "<uploaded frame if supported>" \
      --async \
      --json
    
    genmedia status <endpoint_id> <request_id> \
      --download "./outputs/cinema/{request_id}_{index}.{ext}" \
      --json
    

Prompt build order

Use the SCLCAM structure:

  1. Subject: who or what is in frame.
  2. Context: location, time, weather, story moment.
  3. Lens/framing: distance, angle, focal length feel, depth of field.
  4. Camera motion: only for video or if motion blur is desired.
  5. Atmosphere: haze, rain, practicals, reflections, texture.
  6. Mood/color: palette, contrast, grade, exposure style.
  7. Output controls: aspect ratio, duration, first-frame continuity.

Example structure:

[subject] in [context], framed as [shot size and angle], [lens feel],
[lighting setup], [camera movement if video], [color grade], [texture],
[duration or aspect ratio], [continuity constraints]

Model routing

  • Premium realistic still: use openai/gpt-image-2.
  • Premium stylized still: use openai/gpt-image-2, then fal-ai/nano-banana-pro, then fal-ai/nano-banana-2.
  • Fast draft still: use fal-ai/flux-2/klein/9b.
  • Highest quality video: use bytedance/seedance-2.0/text-to-video or bytedance/seedance-2.0/image-to-video.
  • Motion from a strong frame: use bytedance/seedance-2.0/image-to-video.
  • Fast or lower-cost video: use xai/grok-imagine-video/text-to-video or xai/grok-imagine-video/image-to-video.
  • Complex camera language: inspect Seedance 2.0 first, then Kling v3 when multi-prompt or element controls matter.
  • Story sequence: use the storytelling skill with this skill as shot-language support.
  • Character or product continuity: use the relevant domain skill first, then apply cinematography as the variable block.

Quality bar

Before returning, check:

  • Camera movement is physically plausible for the scene.
  • Lens, shot size, and camera angle do not contradict each other.
  • Lighting direction is clear and consistent.
  • Color grade supports the mood without flattening subject detail.
  • Video prompt describes one shot unless the selected model supports multiple prompts or shot lists.
  • Downloaded files come from downloaded_files[].

If a result looks generic, improve specificity in camera, blocking, light, and environment before adding more adjectives.

Individual skills in this repo

This repo contains 3 individual skills — each has its own dedicated page.

ido6/ai-video-skills

Build Blender previs/blockout scenes in the pawn style (box-torso + sphere-head characters in flat-coloured primitive sets) and export them as motion/structure references for AI video (Seedance, idoGen, Kling).

ido6/ai-video-skills

Elite cinematic-video-prompt discipline for idoGen (mcp__idogen__generate_video, WaveSpeed/fal/poyo Seedance 2.0 and 2.5, Veo 3.1, Kling 3.0). Adapted from a production studio's CINEDANCE prompt-director system, folded together with GenHQ classroom technique (camera-movement library, Seedance 2.0-vs-2.5 dialect split, Kling 3.0 burger + @tag binding, Omni Flash video-to-video lock/preserve-list, storyboard mode, UGC/organic mode). Use this WHENEVER writing or fixing a prompt for mcp__idogen__generate_video, or when the user wants a cinematic/consistent/high-budget-looking video shot, camera control, a specific lens feel, spatial blocking across multiple shots, a scene that must match a previous shot's geography/lighting/character, a named camera move, a Kling multi-shot sequence, an edit of existing footage (Omni-style), or an organic/UGC/phone-shot look. Trigger even without the word "CINEDANCE" — any idoGen video-prompt build or repair task qualifies. Companion skills: idogen-image-brain (references/plat...

ido6/ai-video-skills

Generate Seedance 2.5 prompts for **multishot omni reference mode** — sequences where multiple reference images (@Image1, @Image2, etc.), and optionally videos and audio, are assigned across different shots. Use this skill when the user says "multishot", "omni reference", "omniref", uploads multiple references for a shot sequence, or asks for a multishot Seedance prompt. Do NOT use for single-shot, first+last frame, or generic Seedance prompts — those are separate skills. If the user just says "Seedance prompt" without indicating multishot, ask first.

相關技能