Communitygithub.com

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.

ComfyUI-MiniMax-H3-SkillBridge란 무엇인가요?

ComfyUI-MiniMax-H3-SkillBridge is a Claude Code agent skill that write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.

지원 대상~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge/tree/HEAD/skills/h3-prompt-writing

즐겨 사용하는 AI에게 물어보기

이 에이전트 스킬이 미리 로드된 새 채팅을 엽니다.

문서

H3 Prompt Writing

Workflow

  1. Identify the input mode: T2VA, I2VA, FL2VA, L2VA, or full-reference Ref2VA.
  2. For base text/keyframe modes, read references/base-en.txt and follow its final prompt structure.
  3. For full-reference mode, read references/ref-en.txt and follow its six-section rewrite format.
  4. Preserve the exact field names, section order, labels, and timing notation from the selected guide.

Base Modes

  • T2VA: build the full audiovisual timeline from text.
  • I2VA: start from the first frame and develop forward from it.
  • FL2VA: describe the continuous path between the first and last frames.
  • L2VA: infer a plausible opening and converge to the supplied last frame.

Use integrated_multimodal_description, overall_soundscape, and non_diegetic_music in the order shown in references/base-en.txt.

Full-Reference Mode

Ref2VA rewrites use subject_definitions, summary, retention_analysis, detailed_description, overall_soundscape, and non_diegetic_music in that order. Reference labels stay consistent across all sections.

Read references/ref-en.txt for label rules, retention analysis, and complete examples.

Output Rules

  • Write rewrite sections in English; preserve dialogue, lyrics, and visible scene text in their original language.
  • Describe each shot by composition, subjects, environment, actions, camera, sound, and the exact point where referenced content appears.
  • Avoid plot summaries, unresolved reference labels, and timing that does not match the requested duration.

Individual skills in this repo

This repo contains 10 individual skills — each has its own dedicated page.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

Create complete stylized 3D animated shorts from a story idea through an ordered production workflow covering project brief, story outline, character and environment cards, standardized shot planning, text or optional pencil storyboards, video-model selection, single-shot generation, assembly, BGM matching, and final review. Use when the user wants an end-to-end narrative animation workflow with strong character consistency, scene continuity, timing, camera, performance, and audio control. Not for single images, simple edits, photorealistic live action, or one standalone clip.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

For marketers and creators producing promotional content for brands, products, websites, apps, shops, or personal projects. Users provide logos, product images, interface screenshots, official links, or other verifiable assets and confirm duration, aspect ratio, audience, and campaign focus. The Skill organizes brand facts and asset provenance, selects a narrative direction, plans precise beats and shots, generates needed imagery, video, voiceover, or music, and completes assembly and pre-delivery review. It outputs a promotional short that highlights product capabilities, use cases, and a call to action. Best for launches, website showcases, and social promotion; not for imitating real brand marks without authorized assets, inventing product claims, or producing long-form narrative films.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

For users creating a two-player co-op game menu or opening animation. Users provide two player names, a game title, a target visual style, and optional character reference images. The Skill locks identity cues, generates an approval image from a fixed menu framework with coordinated color, buttons, icons, and typography, then uses the approved result to rebuild the character, UI-copy, and event timing instructions for the final video. It outputs a co-op game intro featuring two characters, player cards, and menu interaction motion. Best for game concepts, character-led menus, and social content; not for playable game development, complex multi-page UI, exact brand-logo replication, or generic character-free title sequences.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

为 MiniMax H3 图生视频生成高动态运镜 Ref2VA 提示词。基于插件自动加载的参考图(人物、场景、动作、风格),只输出可直接交给 H3 图生视频节点的六段式英文提示词,不生成角色卡或动作首帧文生图提示词。要求两轴位移、主动运镜、前景视差、障碍接触、尺度变化与物理反馈,使用官方镜头指令与自然语言弧线描述。无参考图时允许纯文字描述模式。

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

For creators making surreal short videos that blend rough glowing hand-drawn animation with live-action spaces. Users provide a scene idea, contact object or hand, desired mood, and optional language or style constraints. The Skill clarifies the physical contact, designs continuous morphing, escape route, and delayed handheld chase movement, then writes a reusable 15-second 16:9 video prompt in the user's language. After user confirmation it recommends MiniMax H3 generation and checks contact realism, camera delay, rough glowing stroke texture, and non-horror tone. Best for single-scene creative clips, not polished CG, horror jump scares, plush characters, or multi-scene cuts.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

Turn user-provided reference images and script into multi-segment H3 Ref2VA video prompts with character presenter, holographic interaction, and full narration coverage. Users provide a character reference image, an environment reference image, and a complete narration script. The Skill auto-splits the script into segments within user-specified max duration, calculates narration budget dynamically, designs at least 2 holographic interactions per segment, and outputs complete self-contained Ref2VA format prompts. No external skill dependencies. Style, aspect ratio, and segment duration are flexible.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

Turn product images and ad requirements into minimalist product ad shorts for e-commerce promotion and product launches. The Skill confirms format and product variants, extracts selling points, writes concise English ad copy, builds product anchors, plans beat-synced typography/storyboards, and generates a clean product film with premium camera language. Not for KOC talking-head ads, general editing, or complex screen demos.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

For musicians, video creators, and social-media editors producing AI music videos or emotional short films with lyric typography. Users provide music, lyrics, references, characters, typography direction, mood, or target platform. The Skill analyzes beat and vocal timing, separates character, scene, and text references, designs beat-reactive spatial typography, decomposes long works into connected shots, audits prompts, and routes generation for H3 or other video tools. It outputs MV concepts, shot prompts, lyric text plans, and stitching guidance. Best for stylized MVs and subtitle-driven music visuals, not ordinary caption cleanup, licensed IP copying, or fully manual post-production editing.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

For creators, educators, and social-video editors who need a tactile paper-collage language for narration, knowledge points, opinions, or abstract topics. Users provide source copy, story beats, or a core concept and may specify aspect ratio, duration, palette, and audio needs. The Skill extracts meaning, proposes visual metaphors, prepares a production plan and storyboard, generates approved halftone collage stills, then creates stop-motion clips with paper movement and tactile sound effects, with optional final assembly. By default it keeps collage SFX and does not add BGM, voiceover, or subtitles unless requested. Best for explainers, viewpoints, story visuals, and social B-roll; not for presenter ads, editable layers, complex typography, or prompt-only tasks.

StarAI-2026/ComfyUI-MiniMax-H3-SkillBridge

For creators explaining science, education, or general knowledge through tactile handmade papercraft visuals. Users provide a topic, core knowledge points, or source material and may specify audience, duration, aspect ratio, and deliverable type. The Skill extracts the learning goal and visual metaphor, proposes creative directions, designs paper characters, layered diorama sets, and props, creates preview concepts plus image and video prompts, and plans storyboards, camera movement, transitions, and sound with staged approvals and review checklists. It outputs a production-ready papercraft stop-motion explainer package, or selected assets such as still prompts, image-series prompts, short-video prompts, or storyboards. Best for cut-paper, pop-up-book, layered diorama, and miniature stop-motion explainers; not for standard 2D cartoons, line doodles, live action, or explainers without a paper-art look.

관련 스킬