Communitygithub.com

hiapi-product-spokesperson-video

Generate product spokesperson, brand introduction, and on-camera advertising videos locally through HiAPI. Use for synthetic presenters, authorized talking-head videos from a person image, product introductions, or brand promotions from product and brand reference images. Route synthetic spokespersons to Kling 3.0 Omni text-to-video, talking heads to Kling 3.0 Omni image-to-video, and product or brand promos to Seedance 2.0 Fast. Do not use to edit or assemble existing footage.

hiapi-product-spokesperson-video とは?

hiapi-product-spokesperson-video is a Claude Code agent skill that generate product spokesperson, brand introduction, and on-camera advertising videos locally through HiAPI. Use for synthetic presenters, authorized talking-head videos from a person image, product introductions, or brand promotions from product and brand reference images. Route synthetic spokespersons to Kling 3.0 Omni text-to-video, talking heads to Kling 3.0 Omni image-to-video, and product or brand promos to Seedance 2.0 Fast. Do not use to edit or assemble existing footage.

対応~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/HiAPIAI/hiapi-product-video-skills/tree/main/skills/product-spokesperson

お気に入りのAIに質問する

このエージェントスキルを事前に読み込んだ状態で新しいチャットを開きます。

ドキュメント

HiAPI Product Spokesperson Video

Create 3-15 second, single-shot product videos with the local CLI. Always estimate with --dry-run before a paid request. Create a paid task only when the user explicitly approves generation and the command includes --spend.

Route The Scenario

Choose exactly one scenario:

ScenarioModelRequired inputP0 default
synthetic-spokespersonkling-3.0-omni/text-to-videoPrompt and optional dialogue3s, 720p, 9:16, audio
talking-headkling-3.0-omni/image-to-videoOne authorized person image, prompt, dialogue, consent3s, 720p, audio
product-introseedance-2.0-fastPrompt and 1-9 product images4s, 480p, 9:16, audio
brand-promoseedance-2.0-fastPrompt and 1-9 brand or product images4s, 480p, 9:16, audio

Before building the request, confirm the target duration. If the user did not specify it, ask how many seconds the video should be and state the supported range: Kling 3-15 seconds; Seedance 4-15 seconds. Offer the P0 default from the table, but do not proceed to --dry-run until the user chooses a duration or explicitly accepts the default.

Refuse talking-head generation unless the user confirms authorization and the command includes --consent-confirmed. Do not invent product claims, endorsements, prices, certifications, or brand facts.

Protect Spend And Secrets

Run the zero-cost checks first:

node scripts/hiapi-product-spokesperson-video.mjs --check --env-file "D:\path\to\.env.local"

Then run the intended request with --dry-run. Confirm the public-pricing estimate is within the user's limit before replacing --dry-run with --spend. The default client-side estimate limit is $0.50; override it only within an explicit user-approved budget. Create no more than one paid task per invocation.

Treat this as an estimate guard, not a server-enforced final-charge cap. HiAPI may apply an account-specific group ratio during task precharge. Until the task API supports a server-side maximum-cost field, disclose this limitation before paid generation and keep substantial headroom below the user's total budget.

Read HIAPI_API_KEY from the process environment or --env-file. Never print, return, persist, or include the key in previews, artifacts, manifests, errors, or logs. Use temporary storage in P0.

After installation, copy .env.example to an ignored .env, or point --env-file at an existing ignored environment file. Never commit the populated file. Redact both embedded image data and remote image URLs from previews and saved request artifacts.

Run Locally

Estimate an authorized talking-head request without spending:

node scripts/hiapi-product-spokesperson-video.mjs `
  --scenario talking-head `
  --image-file "D:\media\authorized-person.jpg" `
  --prompt "Fixed camera, natural eye contact, restrained movement" `
  --dialogue "New product, now available." `
  --consent-confirmed `
  --env-file "D:\path\to\.env.local" `
  --max-cost-usd 0.50 `
  --dry-run

For synthetic-spokesperson, omit the image and consent options. For product and brand scenarios, repeat --reference-image-file or --reference-image-url. Local JPEG, PNG, and WebP files are converted to data URIs and must not exceed 20 MiB each.

Resume Without Spending Again

The CLI prints the idempotency key before submission and prints the task ID immediately after creation. If submission is interrupted before a task ID is returned, retry with the same --idempotency-key. If polling is interrupted after creation, run only --task-id ID, --env-file, and the desired --output-dir. This resumes polling and downloads the existing result without rebuilding or resubmitting the original request.

Preserve The Output Package

Write each successful run under outputs/ unless --output-dir is supplied. Preserve:

  • final.mp4;
  • script.md and captions.srt;
  • prompt.json with embedded media and remote image URLs redacted;
  • manifest.json with scenario, model, task ID, status, estimate, request hash, storage, timestamp, and expiry;
  • qc.json marked for human review of speech, lip sync, consistency, continuity, and claim accuracy.

Never claim success without a terminal successful task and a usable local video.

Handle Failures

  • Do not retry HTTP 400, 401, 402, or 403.
  • Retry submission HTTP 409, 429, or 503 at most three times with the original idempotency key.
  • Resume polling interruptions with --task-id; never submit a replacement automatically.
  • Do not regenerate after a quality failure without explicit approval for another paid run.
  • Report the task ID, terminal status, estimate, local output path, and pending human QC checks.

Individual skills in this repo

This repo contains 3 individual skills — each has its own dedicated page.

hiapi-fashion-lookbook-video

Plan, preview, price, and generate short single-shot fashion videos through HiAPI. Use for runway walks, one target-outfit transition, editorial lookbook poses, and non-speaking person showcases for clothing brands, fashion creators, and ecommerce sellers. Route text-only briefs to Kling 3.0 Omni text-to-video, one authorized person or complete-look image to Kling image-to-video, and person plus garment/style references to Seedance 2.0 Fast. Do not use for exact virtual try-on or fit guarantees, multi-outfit montages, deterministic motion transfer, talking presenters, captions, logos, CTA compositing, or multi-shot editing.

hiapi-food-commercial-video

Plan, preview, price, and generate single-shot food and beverage commercial videos through HiAPI, with 3-15 second Kling output or 4-15 second Seedance Fast output. Use for coffee pours, beverage splashes, food macro shots, packaged-product hero ads, restaurant atmosphere spots, and ecommerce food promos. Route text-only requests to Kling 3.0 Omni text-to-video, one hero image to Kling image-to-video, and 1-9 reference images to Seedance 2.0 Fast. Do not use for human spokespersons, reference-motion transfer, editing, multi-shot assembly, captions, added logos, or CTA compositing.

hiapi-seedance-2-0-ugc-ad-video-skill

Create source-grounded, real-person or consented synthetic-actor UGC product ads with HiAPI Seedance 2.0 for TikTok and Instagram Reels. Use for ecommerce product trials, unboxings, direct-to-camera demos, creator-style scripts, paid-social variants, 9:16 video generation, or UGC campaign batches that need authentic phone-shot direction, native dialogue, claim controls, and final media QC.

関連スキル