Communitygithub.com

modbender/skill-library-mcp

Generate AI-powered notes from videos (document, outline, or graphic-text formats)

skill-library-mcp 是什麼?

skill-library-mcp is a Claude Code agent skill that generate AI-powered notes from videos (document, outline, or graphic-text formats).

相容平台~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/modbender/skill-library-mcp/tree/HEAD/data/ai-notes-ofvideo

在你喜歡的 AI 中提問

開啟一個已預先載入此 Agent Skill 的新對話。

說明文件

AI Video Notes

Generate structured notes from video URLs using Baidu AI. Supports three note formats.

Workflow

  1. Create Task: Submit video URL → get task ID
  2. Poll Status: Query task every 3-5 seconds until completion
  3. Get Results: Retrieve generated notes when status = 10002

Status Codes

CodeStatusAction
10000In ProgressContinue polling
10002CompletedReturn results
OtherFailedShow error

Note Types

TypeDescription
1Document notes
2Outline notes
3Graphic-text notes

APIs

Create Task

Endpoint: POST /v2/tools/ai_note/task_create

Parameters:

  • video_url (required): Public video URL

Example:

python3 scripts/ai_notes_task_create.py 'https://example.com/video.mp4'

Response:

{
  "task_id": "uuid-string"
}

Query Task

Endpoint: GET /v2/tools/ai_note/query

Parameters:

  • task_id (required): Task ID from create endpoint

Example:

python3 scripts/ai_notes_task_query.py "task-id-here"

Response (Completed):

{
  "status": 10002,
  "notes": [
    {
      "tpl_no": "1",
      "contents: ["Note content..."]
    }
  ]
}

Polling Strategy

Option 1: Manual Polling

  1. Create task → store task_id
  2. Query every 3-5 seconds:
    python3 scripts/ai_notes_task_query.py <task_id>
    
  3. Show progress updates:
    • Status 10000: Processing...
    • Status 10002: Completed
  4. Stop after 30-60 seconds (video length dependent)

Option 2: Auto Polling (Recommended)

Use the polling script for automatic status updates:

python3 scripts/ai_notes_poll.py <task_id> [max_attempts] [interval_seconds]

Examples:

# Default: 20 attempts, 3-second intervals
python3 scripts/ai_notes_poll.py "task-id-here"

# Custom: 30 attempts, 5-second intervals
python3 scripts/ai_notes_poll.py "task-id-here" 30 5

Output:

  • Shows real-time progress: [1/20] Processing... 25%
  • Auto-stops when complete
  • Returns formatted notes with type labels

Error Handling

  • Invalid URL: "Video URL not accessible"
  • Processing error: "Failed to parse video"
  • Timeout: "Video too long, try again later"

Individual skills in this repo

This repo contains 12 individual skills — each has its own dedicated page.

modbender/skill-library-mcp

Native video analysis using Google Gemini API. Upload and analyze video files — describe scenes, extract text/UI, answer questions about content, transcribe speech, identify objects and actions. Use when: (1) User sends a video file and wants it analyzed, (2) Video summarization or description needed, (3) Extracting text, UI elements, or information from screen recordings, (4) Answering questions about video content, (5) Comparing multiple videos, (6) Analyzing tutorials, demos, or walkthroughs.

modbender/skill-library-mcp

Borrow from Aave via credit delegation. Agent self-funds by borrowing against delegator collateral. Supports borrow, repay, health checks. Works on Aave V2/V3.

modbender/skill-library-mcp

Create AI avatar and talking head videos with OmniHuman, Fabric, PixVerse via inference.sh CLI. Models: OmniHuman 1.5, OmniHuman 1.0, Fabric 1.0, PixVerse Lipsync. Capabilities: audio-driven avatars, lipsync videos, talking head generation, virtual presenters. Use for: AI presenters, explainer videos, virtual influencers, dubbing, marketing videos. Triggers: ai avatar, talking head, lipsync, avatar video, virtual presenter, ai spokesperson, audio driven video, heygen alternative, synthesia alternative, talking avatar, lip sync, video avatar, ai presenter, digital human

modbender/skill-library-mcp

AI代码编辑器规则模板集合 - 为Cursor、Windsurf、Claude Code、Cline等AI编辑器提供项目规则配置。适用于需要配置AI编码助手规则的开发者,包含全栈Web、移动端、Vue3+SpringBoot等技术栈模板。

modbender/skill-library-mcp

Generate landing page code from product descriptions

modbender/skill-library-mcp

Create AI marketing videos for ads, promos, product launches, and brand content. Models: Veo, Seedance, Wan, FLUX for visuals, Kokoro for voiceover. Types: product demos, testimonials, explainers, social ads, brand videos. Use for: Facebook ads, YouTube ads, product launches, brand awareness. Triggers: marketing video, ad video, promo video, commercial, brand video, product video, explainer video, ad creative, video ad, facebook ad video, youtube ad, instagram ad, tiktok ad, promotional video, launch video

modbender/skill-library-mcp

Generate AI music videos end-to-end. Creates music with Suno (sunoapi.org), generates visuals with OpenAI/Seedream/Google/Seedance, and assembles into music video with ffmpeg. Supports timestamped lyrics (auto SRT), Suno native music video generation, slideshow/video/hybrid modes. Token-based cost tracking per generation.

modbender/skill-library-mcp

The video AI notes tool is provided by Baidu. Based on the video download address provided by the user, it downloads and parses the video, and finally generates AI notes corresponding to the video (a total of three types of notes can be generated: document notes, outline notes, and image-text notes).

modbender/skill-library-mcp

Lead strategy with competitive analysis, strategic planning, growth roadmaps, and business portfolio management.

modbender/skill-library-mcp

Build high-converting landing pages with proven structure, persuasive copy, and conversion optimization.

modbender/skill-library-mcp

Web SaaS service decomposition and AI internalization development plan generation. Analyze existing SaaS to identify functions replaceable by AI agents and establish skill-based internalization roadmap. Triggered by 'SaaS analysis', 'service decomposition', 'internalization', 'decompose', 'SaaS replacement', 'build this service with AI', etc.

modbender/skill-library-mcp

Long-form AI video production: the frontier of multi-agent coordination. CellCog orchestrates 6-7 foundation models to produce up to 4-minute videos from a single prompt — scripted, filmed, voiced, lipsync'd, scored, and edited automatically. Create marketing videos, product demos, explainer videos, educational content, spokesperson videos, training materials, UGC content, news reports.

相關技能