Communitygithub.com

curiositech/windags-skills

Expert in script-to-video production pipelines for Apple Silicon Macs. Specializes in hybrid local/cloud workflows, LoRA training for character consistency, motion graphics generation, and artist commissioning. Activate on 'AI video production', 'script to video', 'video generation pipeline', 'character consistency', 'LoRA training', 'cloud GPU', 'motion graphics', 'Wan I2V', 'InVideo alternative'. NOT for real-time video editing, video compositing (use DaVinci/Premiere), audio production, or 3D modeling (use Blender/Maya).

¿Qué es windags-skills?

windags-skills is a Claude Code agent skill that expert in script-to-video production pipelines for Apple Silicon Macs. Specializes in hybrid local/cloud workflows, LoRA training for character consistency, motion graphics generation, and artist commissioning. Activate on 'AI video production', 'script to video', 'video generation pipeline', 'character consistency', 'LoRA training', 'cloud GPU', 'motion graphics', 'Wan I2V', 'InVideo alternative'. NOT for real-time video editing, video compositing (use DaVinci/Premiere), audio production, or 3D modeling (use Blender/Maya).

Compatible con~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/curiositech/windags-skills/tree/HEAD/skills/ai-video-production-master

Preguntar en tu IA favorita

Abre un nuevo chat con esta habilidad de agente ya precargada.

Documentación

AI Video Production Master

Expert in script-to-video production pipelines for Apple Silicon Macs. Specializes in hybrid local/cloud workflows, style consistency, and motion graphics generation.

DECISION POINTS

Method Selection Tree

Script Analysis:
├── Content Type = Educational/Corporate/Documentary
│   ├── Budget < $50/month → Stock Footage Assembly (InVideo-style)
│   └── Budget > $50/month → Stock + T2V Hybrid
├── Content Type = Creative/Artistic/Abstract
│   ├── Timeline < 2 days → Sora/Runway T2V
│   └── Timeline > 2 days → Custom I2V with LoRA training
└── Content Type = Brand/Character-focused
    ├── Existing assets available → I2V animation pipeline
    └── No assets → Commission artists → I2V pipeline

Quality vs. Cost Decision:
├── Professional deliverable needed
│   ├── Use Sora/Runway Gen-3 (premium T2V)
│   └── Commission custom artwork for I2V
├── Prototype/test content
│   ├── Use stock footage + free T2V models
│   └── Local ComfyUI I2V generation
└── High volume production
    ├── Cloud batch processing (Vast.ai)
    └── Automated stock footage workflows

Error Recovery Tree:
├── Quality fails at T2V generation
│   ├── Switch to stock footage + motion graphics overlay
│   └── Fallback to I2V with commissioned artwork
├── Cloud GPU timeout/failure
│   ├── Retry with different provider (Vast.ai → RunPod)
│   └── Switch to local processing with extended timeline
└── Style consistency breaks
    ├── Retrain LoRA with more reference images
    └── Use IPAdapter + consistent prompt structure

Processing Location Decision Matrix

Shot CountComplexityBudgetRecommendation
1-5SimpleAnyLocal M4 Max
6-20Medium<$10Stock footage
6-20High>$10Cloud GPU batch
20+AnyAnyCloud GPU required

FAILURE MODES

Blurry Motion Jitter

Symptoms: Inconsistent frame rates, stuttering motion, temporal artifacts Diagnosis: Mismatched FPS settings or insufficient motion guidance Fix: Set consistent 24fps pipeline, add motion strength controls, use higher guidance scale (7-9) Detection Rule: If temporal consistency score < 0.7 or frame delta > 0.3, apply motion stabilization

Character Consistency Drift

Symptoms: Character appearance changes between shots, style inconsistency Diagnosis: Insufficient reference conditioning or LoRA overfitting Fix: Retrain LoRA with 15-20 reference images, use IPAdapter for face consistency Detection Rule: If character similarity score < 0.8 between consecutive shots, halt and retrain

Cost Overrun Spiral

Symptoms: Cloud GPU bills exceed budget by >200%, slow iteration cycles Diagnosis: No cost monitoring or inefficient batch sizing Fix: Set hard limits in cloud scripts, optimize batch sizes, use spot instances Detection Rule: If cost-per-minute exceeds $0.50 or iteration time > 15min, switch to local processing

Audio Desync Cascade

Symptoms: Lip sync drift, audio-visual timing mismatches Diagnosis: Variable generation times affecting audio alignment Fix: Generate all video first, then align audio in post with markers Detection Rule: If audio offset > 200ms from video markers, re-align with FFmpeg

Schema Bloat Paralysis

Symptoms: Too many generation options, analysis paralysis, project stalls Diagnosis: Over-optimization without testing simple approaches first Fix: Always start with stock footage proof-of-concept, iterate to complexity Detection Rule: If project planning > 2 hours without generated content, default to stock footage

WORKED EXAMPLES

Educational Video: "Quantum Computing Explained"

Input: 3-minute script about quantum computing basics

Decision Process:

  1. Content analysis: Educational → Stock footage recommended
  2. Budget check: $30 budget → Use Pexels API + motion graphics
  3. Shot breakdown: 12 shots needed
  4. Quality target: Professional but not cinematic

Execution:

# Step 1: Generate shot list from script
python scripts/script_analyzer.py --script quantum_script.txt
# Output: 12 shots identified, 8 require stock footage, 4 need motion graphics

# Step 2: Source stock footage
python scripts/stock_video_generator.py --shots_file shots.json --style documentary
# Selected: Laboratory footage, abstract tech visuals, clean backgrounds

# Step 3: Generate motion graphics for complex concepts
python scripts/motion_graphics_generator.py --type data_viz --concept "quantum_states"
# Created: Animated diagrams for superposition, entanglement

# Step 4: Assembly and sync
python scripts/video_assembler.py --footage stock_clips/ --graphics motion/ --audio narration.wav

Expert vs. Novice:

  • Expert: Checked audio levels before graphics generation, chose 16:9 format for platform compatibility
  • Novice would miss: Consistent lighting in stock selection, proper motion graphic duration matching speech pace

Final output: 3:15 video, cost $0 (free stock), 45 minutes total production time

QUALITY GATES

Video production complete when ALL conditions met:

  • Frame rate consistent at 24fps across all segments
  • Audio levels normalized between -23dB and -18dB LUFS
  • Visual style maintains consistency (color grading, aspect ratio)
  • Motion graphics sync with narration within 100ms tolerance
  • Character/subject consistency score >0.8 across shots
  • No temporal artifacts (flicker, jitter) in any 5-second segment
  • Subtitle/caption timing aligned with audio peaks
  • Export renders without errors in target resolution
  • File size within platform limits (YouTube: <128GB, social: <4GB)
  • Backup files stored with version control

NOT-FOR BOUNDARIES

This skill should NOT be used for:

  • Real-time video editing → Use DaVinci Resolve or Premiere Pro directly
  • Complex video effects/compositing → Delegate to After Effects or physics-rendering-expert skill
  • Audio production beyond basic sync → Use sound-engineer skill for mixing, mastering, sound design
  • 3D modeling or complex animation → Use Blender/Maya or physics-rendering-expert skill
  • Live streaming or broadcast → Use OBS Studio or broadcast-specific tools
  • Color grading beyond basic correction → Use professional colorist workflows
  • Motion tracking or match moving → Use specialized tracking software
  • Multi-camera synchronization → Use dedicated sync tools like PluralEyes

Handoff triggers:

  • If project requires >20 layers of compositing → physics-rendering-expert
  • If audio needs custom sound design → sound-engineer
  • If 3D elements need modeling → Blender specialist
  • If real-time performance needed → Use native video editing software

Individual skills in this repo

This repo contains 20 individual skills — each has its own dedicated page.

curiositech/windags-skills

Expert in 2000s-era music visualization (Milkdrop, AVS, Geiss) and modern WebGL implementations. Specializes in Butterchurn integration, Web Audio API AnalyserNode FFT data, GLSL shaders for audio-reactive visuals, and psychedelic generative art. Activate on "Milkdrop", "music visualization", "WebGL visualizer", "Butterchurn", "audio reactive", "FFT visualization", "spectrum analyzer". NOT for simple bar charts/waveforms (use basic canvas), video editing, or non-audio visuals.

curiositech/windags-skills

Expert legal research agent for finding and scraping expungement data state by state. Knows authoritative sources, URL patterns, Firecrawl configuration, and 2026 legal landscape.

curiositech/windags-skills

Expert in 3D computer vision labeling tools, workflows, and AI-assisted annotation for LiDAR, point clouds, and sensor fusion. Covers SAM4D/Point-SAM, human-in-the-loop architectures, and vertical-specific training strategies. Activate on '3D labeling', 'point cloud annotation', 'LiDAR labeling', 'SAM 3D', 'SAM4D', 'sensor fusion annotation', '3D bounding box', 'semantic segmentation point cloud'. NOT for 2D image labeling (use clip-aware-embeddings), general ML training (use ml-engineer), video annotation without 3D (use computer-vision-pipeline), or VLM prompt engineering (use prompt-engineer).

curiositech/windags-skills

Implement WCAG 2.2 AA/AAA compliance with automated testing, keyboard navigation, screen reader support, and focus management. Activate on: accessibility audit, WCAG compliance, keyboard navigation, screen reader, aria attributes, axe-core, focus trap. NOT for: design-level accessibility review (use design-accessibility-auditor), color contrast only (use css-in-js-architect).

curiositech/windags-skills

Time-blind friendly planning, executive function support, and daily structure for ADHD brains. Specializes in realistic time estimation, dopamine-aware task design, and building systems that actually work for neurodivergent minds.

curiositech/windags-skills

Designs digital experiences for ADHD brains using neuroscience research and UX principles. Expert in reducing cognitive load, time blindness solutions, dopamine-driven engagement, and compassionate design patterns. Activate on 'ADHD design', 'cognitive load', 'accessibility', 'neurodivergent UX', 'time blindness', 'dopamine-driven', 'executive function'. NOT for general accessibility (WCAG only), neurotypical UX design, or simple UI styling without ADHD context.

curiositech/windags-skills

>- Apply crisis decision-making research to agent routing, uncertainty triage, and coordination failure analysis in time-pressured systems. Use when diagnosing handoff failures, analytical paralysis, or expert judgment under incomplete information. NOT for routine coding, simple CRUD design, or static single-agent tasks with complete information.

curiositech/windags-skills

Extend and modify the admin dashboard, developer portal, and operations console. Use when adding new admin tabs, metrics, monitoring features, or internal tools. Activates for dashboard development, analytics, user management, and internal tooling.

curiositech/windags-skills

Conversation patterns and interaction protocols for multi-agent systems. Covers request/response, pub/sub, blackboard, delegation chains, debate, critique, consensus, fan-out/fan-in, supervisor-worker, and peer negotiation. Deep analysis of AutoGen conversation patterns, CrewAI delegation, LangGraph state passing, and FIPA-ACL performatives. Teaches how to design what agents say to each other and in what order. Activate on: "agent conversation", "agent protocol", "multi-agent debate", "agent delegation", "supervisor worker pattern", "agent voting", "consensus protocol", "fan-out fan-in", "agent negotiation", "blackboard pattern", "agent dialogue", "conversation topology", "agent handoff". NOT for: wire format or serialization (use agent-interchange-formats), orchestration infrastructure (use agentic-infrastructure-2026), single agent behavior (use agentic-patterns).

curiositech/windags-skills

Meta-agent for creating new custom agents, skills, and MCP integrations. Expert in agent design, MCP development, skill architecture, and rapid prototyping. Activate on 'create agent', 'new skill', 'MCP server', 'custom tool', 'agent design'. NOT for using existing agents (invoke them directly), general coding (use language-specific skills), or infrastructure setup (use deployment-engineer).

curiositech/windags-skills

AI-powered calendar management and agent-based scheduling coordination. Covers calendar APIs (Google Calendar, CalDAV/iCal), AI scheduling assistants (Reclaim, Clockwise, Motion, Cal.com), building custom calendar agents with MCP, multi-calendar merging, timezone management, focus block protection, meeting fatigue detection, and agent-to-agent meeting negotiation protocols. Activate on: "calendar agent", "AI scheduling", "calendar coordination", "meeting scheduling", "calendar API", "focus time protection", "calendar optimization", "Google Calendar MCP", "Reclaim", "Clockwise", "Motion", "Cal.com", "smart scheduling", "calendar-aware agent", "timezone scheduling", "agent negotiation meetings". NOT for: manual calendar UI component design (use form-validation-architect), project management scheduling or Gantt charts (use project-management-guru-adhd), general time-tracking or pomodoro apps (use adhd-daily-planner for time-awareness), building the agent itself from scratch (use agent-creator).

curiositech/windags-skills

Build and adopt production AI agent infrastructure in 2026. Covers framework selection (LangGraph, CrewAI, AutoGen, MCP), orchestration patterns, evaluation, observability, memory systems, and tool use. Also covers the SOCIAL dimension: how to sell agent infrastructure internally, change management, measuring ROI, building trust in autonomous systems, and scaling adoption across teams. Activate on: "agent infrastructure", "agent framework comparison", "which agent framework", "sell AI tools internally", "agent adoption", "agent observability", "agent evaluation", "MCP architecture", "agentic mesh", "enterprise AI agents", "AI change management", "agent ROI". NOT for: building specific agents (use ai-engineer), designing agent behavior patterns (use agentic-patterns), prompt tuning (use prompt-engineer).

curiositech/windags-skills

Fundamental patterns for effective agentic behavior. Teaches decomposition, tool orchestration, error recovery, context management, quality self-assessment, and knowing when to stop. Model-agnostic principles that make any agent more effective regardless of domain. Activate on: "how should I structure this agent", "agentic workflow", "agent patterns", "multi-step task", "tool orchestration", "/agentic-patterns", "decompose this", "agent best practices", "chain of actions", "when should the agent stop", "agent loop design". NOT for: creating agent infrastructure (use agent-creator), building DAGs (use windags-architect), specific tool implementation.

curiositech/windags-skills

Automated discovery and matching of agent skills for dynamic task routing and capability assessment

curiositech/windags-skills

Cryptographic security for agentic systems — zero-trust agent networking, signed message envelopes (JWS/JWE), capability-based security (ocaps), Merkle tree audit trails, WASM sandboxing, and formal verification. Covers CLI dev tool security, mTLS between agents, permission boundaries (least privilege for AI agents), and supply chain security for skills/plugins. Activate on: "agent security", "zero trust agents", "secure agent communication", "capability-based security", "ocap", "signed messages between agents", "agent audit trail", "sandbox agent execution", "agent permissions", "mTLS agents", "cryptographic verification", "agent supply chain", "OWASP agentic", "prove agent did X", "tamper-proof agent logs". NOT for: application-level SAST scanning (use security-auditor), network firewall rules (use infrastructure), SOC2/HIPAA compliance (organizational), or prompt injection defense (use prompt-engineer).

curiositech/windags-skills

Data structures and serialization formats for agent-to-agent communication. Covers message envelopes, structured output schemas, capability declarations, task handoff payloads, error/retry signaling, and context windows as data structures. Deep comparison of A2A protocol, MCP, OpenAI function calling, and LangChain message types. Teaches when to use rigid schemas vs free-form with validation, typed vs untyped, streaming vs batch. Activate on: "agent message format", "agent communication schema", "agent-to-agent protocol", "A2A protocol", "MCP message format", "structured output for agents", "agent interop", "interchange format", "agent serialization", "task handoff format", "capability declaration". NOT for: what agents say to each other (use agent-conversation-protocols), orchestration topology (use multi-agent-coordination), building agent infrastructure (use agentic-infrastructure-2026).

curiositech/windags-skills

Logic-based agent programming language implementing BDI architecture for practical autonomous agent development

curiositech/windags-skills

>- Design AgentSpeak(L)-style BDI agents with context-guarded plans, selection functions, and intention stacks. Use for interruptible autonomy, agent policy, and multi-agent orchestration in dynamic environments. NOT for simple rule engines, static planners, or centralized workflows.

curiositech/windags-skills

Foundational concurrent computation model where actors communicate exclusively through asynchronous message passing

curiositech/windags-skills

license: Apache-2.0 NOT for unrelated tasks outside this domain.

Skills relacionados