Communitygithub.com

Aadi-110i/PEP-PROJECT

Best practices and techniques for writing effective AI video generation prompts. Covers: Veo, Seedance, Wan, Grok, Kling, Runway, Pika, Sora prompting strategies. Learn: shot types, camera movements, lighting, pacing, style keywords, negative prompts. Use for: improving video quality, getting consistent results, professional video prompts. Triggers: video prompt, how to prompt video, veo prompts, video generation tips, better ai video, video prompt engineering, video prompt guide, video prompt template, ai video tips, video prompt best practices, video prompt examples, cinematography prompts

PEP-PROJECT 是什么?

PEP-PROJECT is a Claude Code agent skill that best practices and techniques for writing effective AI video generation prompts. Covers: Veo, Seedance, Wan, Grok, Kling, Runway, Pika, Sora prompting strategies. Learn: shot types, camera movements, lighting, pacing, style keywords, negative prompts. Use for: improving video quality, getting consistent results, professional video prompts. Triggers: video prompt, how to prompt video, veo prompts, video generation tips, better ai video, video prompt engineering, video prompt guide, video prompt template, ai video tips, video prompt best practices, video prompt examples, cinematography prompts.

兼容平台~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/Aadi-110i/PEP-PROJECT/tree/HEAD/skills/video-prompting-guide

在你喜欢的 AI 中提问

打开一个已预加载此 Agent Skill 的新对话。

预览

来自技能 README

Video Prompting Guide

文档

Video Prompting Guide

Best practices for writing effective AI video generation prompts via inference.sh.

Video Prompting Guide

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Well-structured video prompt
belt app run google/veo-3-1-fast --input '{
  "prompt": "Cinematic tracking shot of a red sports car driving through Tokyo at night, neon lights reflecting on wet streets, rain falling, 4K, shallow depth of field"
}'

Prompt Structure Formula

[Shot Type] + [Subject] + [Action] + [Setting] + [Lighting] + [Style] + [Technical]

Example Breakdown

"Slow motion close-up of coffee being poured into a white ceramic cup,
steam rising, morning sunlight streaming through window, warm color grading,
cinematic, 4K, shallow depth of field"
  • Shot Type: Slow motion close-up
  • Subject: Coffee
  • Action: Being poured
  • Setting: White ceramic cup, window
  • Lighting: Morning sunlight
  • Style: Warm color grading, cinematic
  • Technical: 4K, shallow depth of field

Shot Types

Shot TypeDescriptionUse For
Wide shotShows entire sceneEstablishing location
Medium shotWaist-up framingConversations, actions
Close-upFace or detailEmotion, product detail
Extreme close-upSingle featureDrama, texture
Aerial shotBird's eye viewLandscapes, scale
Low angleCamera looking upPower, grandeur
High angleCamera looking downVulnerability
Dutch angleTilted cameraUnease, tension
POV shotFirst person viewImmersion

Camera Movements

MovementDescriptionEffect
Tracking shotCamera follows subjectDynamic, engaging
Dolly in/outCamera moves toward/awayFocus, reveal
PanHorizontal rotationSurvey scene
TiltVertical rotationReveal height
Crane shotVertical + horizontalDramatic reveal
HandheldSlight shakeRealism, urgency
SteadicamSmooth followingProfessional, cinematic
ZoomLens zoom in/outQuick focus change
StaticNo movementContemplation, stability

Lighting Keywords

KeywordEffect
Golden hourWarm, soft, romantic
Blue hourCool, moody, twilight
High keyBright, minimal shadows
Low keyDark, dramatic shadows
Rim lightingSubject outlined with light
BacklitLight from behind subject
Soft lightingGentle, flattering
Hard lightingSharp shadows, contrast
NeonColorful, urban, cyberpunk
Natural lightingRealistic, documentary

Style Keywords

Cinematic Styles

cinematic, film grain, anamorphic lens, letterbox,
shallow depth of field, bokeh, 35mm film,
color grading, theatrical

Visual Aesthetics

minimalist, maximalist, vintage, retro, futuristic,
cyberpunk, steampunk, noir, pastel, vibrant,
muted colors, high contrast, desaturated

Quality Keywords

4K, 8K, high resolution, photorealistic,
hyperrealistic, ultra detailed, professional,
broadcast quality, HDR

Prompt Examples by Use Case

Product Demo

belt app run google/veo-3-1-fast --input '{
  "prompt": "Smooth tracking shot around a sleek smartphone on a white pedestal, soft studio lighting, product photography style, reflections on surface, 4K, shallow depth of field"
}'

Nature Documentary

belt app run google/veo-3-1 --input '{
  "prompt": "Slow motion extreme close-up of a hummingbird hovering at a red flower, wings in motion blur, shallow depth of field, golden hour lighting, National Geographic style"
}'

Urban Lifestyle

belt app run google/veo-3 --input '{
  "prompt": "Tracking shot following a cyclist through busy city streets, morning rush hour, natural lighting, handheld camera feel, documentary style, authentic and candid"
}'

Food Content

belt app run bytedance/seedance-1-5-pro --input '{
  "prompt": "Close-up of chocolate sauce being drizzled over ice cream, slow motion, steam rising, soft lighting, food photography style, appetizing, commercial quality"
}'

Tech/Futuristic

belt app run xai/grok-imagine-video --input '{
  "prompt": "Futuristic control room with holographic displays, camera slowly pans across the space, blue and cyan lighting, sci-fi atmosphere, Blade Runner aesthetic, 4K",
  "duration": 5
}'

Common Mistakes to Avoid

MistakeProblemBetter Approach
Too vague"A nice video"Specify shot, subject, style
Too complexMultiple scenesOne scene per prompt
No motionStatic descriptionInclude camera movement or action
Conflicting styles"Minimalist maximalist"Choose one aesthetic
No lightingUndefined moodSpecify lighting conditions

Model-Specific Tips

Google Veo

  • Excels at realistic, cinematic content
  • Supports audio generation (Veo 3+)
  • Best with detailed, professional prompts
  • Frame interpolation available in 3.1

Seedance

  • Strong at dance and human motion
  • First-frame control available
  • Good for consistent character motion
  • Works well with reference images

Wan 2.5

  • Best for image-to-video
  • Animates still images naturally
  • Good motion prediction
  • Works with any image style

Grok

  • Good general-purpose video
  • Configurable duration (5-10s)
  • Creative interpretations
  • Works well with abstract concepts

Workflow: Iterative Prompting

# 1. Start with basic prompt
belt app run google/veo-3-1-fast --input '{
  "prompt": "A woman walking through a forest"
}'

# 2. Add specificity
belt app run google/veo-3-1-fast --input '{
  "prompt": "Medium tracking shot of a woman in a red dress walking through an autumn forest"
}'

# 3. Add style and technical details
belt app run google/veo-3-1-fast --input '{
  "prompt": "Cinematic medium tracking shot of a woman in a flowing red dress walking through an autumn forest, golden hour sunlight filtering through leaves, shallow depth of field, film grain, 4K"
}'

Related Skills

# Generate videos
npx skills add inference-sh/skills@ai-video-generation

# Google Veo specific
npx skills add inference-sh/skills@google-veo

# Generate images for image-to-video
npx skills add inference-sh/skills@ai-image-generation

# General prompt engineering
npx skills add inference-sh/skills@prompt-engineering

# Full platform skill
npx skills add inference-sh/skills@infsh-cli

Browse all video apps: belt app list --category video

Documentation

Individual skills in this repo

This repo contains 13 individual skills — each has its own dedicated page.

Aadi-110i/PEP-PROJECT

Create AI avatar and talking head videos via inference.sh CLI. Recommended: P-Video-Avatar (fastest, cheapest, built-in TTS). Also: OmniHuman, Fabric, PixVerse. Capabilities: audio-driven avatars, text-to-avatar, lipsync videos, talking head generation, virtual presenters. Use for: AI presenters, explainer videos, virtual influencers, dubbing, marketing videos. Triggers: ai avatar, talking head, lipsync, avatar video, virtual presenter, ai spokesperson, audio driven video, heygen alternative, synthesia alternative, talking avatar, lip sync, video avatar, ai presenter, digital human

Aadi-110i/PEP-PROJECT

Create AI marketing videos for ads, promos, product launches, and brand content. Models: Veo, Seedance, Wan, FLUX for visuals, Kokoro for voiceover. Types: product demos, testimonials, explainers, social ads, brand videos. Use for: Facebook ads, YouTube ads, product launches, brand awareness. Triggers: marketing video, ad video, promo video, commercial, brand video, product video, explainer video, ad creative, video ad, facebook ad video, youtube ad, instagram ad, tiktok ad, promotional video, launch video

Aadi-110i/PEP-PROJECT

Generate AI videos with Google Veo, Seedance 2.0, HappyHorse, Wan, Grok and 40+ models via inference.sh CLI. Models: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Capabilities: text-to-video, image-to-video, reference-to-video, video editing, lipsync, avatar animation, video upscaling, foley sound. Use for: social media videos, marketing content, explainer videos, product demos, AI avatars. Triggers: video generation, ai video, text to video, image to video, veo, animate image, video from image, ai animation, video generator, generate video, t2v, i2v, ai video maker, create video with ai, runway alternative, pika alternative, sora alternative, kling alternative, seedance, happyhorse

Aadi-110i/PEP-PROJECT

ElevenLabs text-to-speech with 22+ premium voices, multilingual support, and voice tuning via inference.sh CLI. Models: eleven_multilingual_v2 (highest quality), eleven_turbo_v2_5 (low latency), eleven_flash_v2_5 (ultra-fast). Capabilities: text-to-speech, voice selection, stability/style control, 32 languages. Use for: voiceovers, audiobooks, video narration, podcasts, accessibility, IVR. Triggers: elevenlabs, eleven labs, elevenlabs tts, premium tts, professional voice, ai voice, high quality tts, multilingual tts, eleven labs voice, voice generation, natural speech, realistic voice, voice over, speech synthesis

Aadi-110i/PEP-PROJECT

Explainer video production guide: scripting, voiceover, visuals, and assembly. Covers script formulas, pacing rules, scene planning, and multi-tool pipelines. Use for: product demos, how-it-works videos, onboarding videos, social explainers. Triggers: explainer video, how to make explainer, product video, demo video, video production, video script, animated explainer, product demo video, tutorial video, onboarding video, walkthrough video, video pipeline

Aadi-110i/PEP-PROJECT

Still-to-video conversion guide: model selection, motion prompting, and camera movement. Covers Wan 2.5 i2v, Seedance, Fabric, Grok Video with when to use each. Use for: animating images, creating video from stills, adding motion, product animations. Triggers: image to video, i2v, animate image, still to video, add motion to image, image animation, photo to video, animate still, wan i2v, image2video, bring image to life, animate photo, motion from image

Aadi-110i/PEP-PROJECT

Landing page conversion optimization with layout rules, hero section design, and CTA psychology. Covers above-the-fold formula, social proof placement, mobile design, and F-pattern reading. Use for: startup landing pages, product pages, SaaS marketing, conversion optimization. Triggers: landing page, hero section, above the fold, conversion optimization, landing page design, cta button, hero image, landing page layout, saas landing page, product page design, conversion rate, landing page best practices

Aadi-110i/PEP-PROJECT

Generate talking head avatar videos with Pruna P-Video-Avatar via inference.sh CLI. Turn a portrait image into a realistic speaking video with built-in TTS. 18x faster and 6x cheaper than competitors. Models: P-Video-Avatar, P-Image (for portrait generation). Capabilities: text-to-avatar, audio-driven avatars, 30 voices, 10 languages, 720p/1080p, built-in TTS, dynamic backgrounds, full-body control. Use for: AI presenters, product demos, explainer videos, virtual influencers, marketing, education, multilingual content, UGC, gaming avatars. Triggers: avatar video, talking head, ai avatar, p-video-avatar, pruna avatar, video avatar, ai presenter, digital human, virtual presenter, lipsync, talking avatar, ai spokesperson, heygen alternative, synthesia alternative, veed alternative, fabric alternative, omnihuman alternative

Aadi-110i/PEP-PROJECT

Generate videos with Pruna P-Video and WAN models via inference.sh CLI. Models: P-Video, WAN-T2V, WAN-I2V. Capabilities: text-to-video, image-to-video, audio support, 720p/1080p, fast inference. Pruna optimizes models for speed without quality loss. Triggers: pruna video, p-video, pruna ai video, fast video generation, optimized video, wan t2v, wan i2v, economic video generation, cheap video generation, pruna text to video, pruna image to video

Aadi-110i/PEP-PROJECT

Render videos from React/Remotion component code via inference.sh. Pass TSX code, get MP4. Supports all Remotion APIs: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence. Configurable resolution, FPS, duration, codec. Use for: programmatic video generation, animated graphics, motion design, data-driven videos, React animations to video. Triggers: remotion, render video from code, tsx to video, react video, programmatic video, remotion render, code to video, animated video, motion graphics code, react animation video

Aadi-110i/PEP-PROJECT

Three.js animation - keyframe animation, skeletal animation, morph targets, animation mixing. Use when animating objects, playing GLTF animations, creating procedural motion, or blending animations.

Aadi-110i/PEP-PROJECT

Video ad creation with exact platform-specific specs for TikTok, Instagram, YouTube, Facebook, LinkedIn. Covers dimensions, duration limits, AIDA framework, and caption requirements. Use for: video ads, social media ads, paid media creative, video marketing, ad production. Triggers: video ad, social media ad, tiktok ad, instagram ad, youtube ad, facebook ad, linkedin ad, video creative, ad specs, paid media, video marketing, ad production, reels ad, stories ad, pre roll, bumper ad

Aadi-110i/PEP-PROJECT

YouTube thumbnail design with specific dimensions, contrast rules, and mobile preview optimization. Covers safe zones, text placement, face expression psychology, and A/B testing. Use for: YouTube thumbnails, video cover images, click-through optimization. Triggers: youtube thumbnail, thumbnail design, video thumbnail, click through rate, ctr optimization, youtube cover, video cover image, thumbnail maker, thumbnail tips, youtube design, video preview image

相关技能