Token导航 LogoToken导航TokenDH.com
开发敏感数据clawhub未标认证来源可访问clear审计提醒

xpilot-ad-makerxpilot 广告制作工具

Agent Skill

xpilot-ad-maker 用于辅助视频、动画、脚本化剪辑和多媒体生成流程,适合在 OpenClaw 中需要整理视频素材、生成脚本或维护合成项目时使用。可结合来源仓库、安装命令和原始 README 继续核验具体用法。安装前建议确认权限范围、维护状态,以及是否会触发联网、命令执行或文件读写。

总安装

2,472

周安装

103

GitHub Stars

公开资料未说明

下载量

824
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:xpilot-ad-maker(xpilot 广告制作工具)
来源仓库:https://github.com/jytech2023/xpilot-ad-maker
安装命令:
openclaw skills install xpilot-ad-maker
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install xpilot-ad-maker

简介

xpilot-ad-maker 用于辅助视频生成流程,适合在 OpenClaw 中需要制作广告视频时使用。

  • 适用于生成 30 秒电影广告视频的场景。
  • 核心能力包括角色一致性、AI 旁白和品牌叠加功能。
  • 通过 clawhub 安装,命令为 openclaw skills install xpilot-ad-maker。
  • 安装前需确认权限范围和维护状态,注意可能触发外部 API 调用操作。

SKILL.md

name
xpilot-ad-maker
description
Generate a 30-second cinematic ad video with consistent character, AI narration, brand overlays, and ambient music. Uses Vidu reference-to-video for character continuity across 4 scenes. Demoed on a medical-tourism storyboard but works for any vertical.
license
MIT-0
metadata
openclaw
version
0.1.0
emoji
🎬
homepage
https://github.com/dotku/x-post-scheduler/tree/main/skills/xpilot-ad-maker
primaryEnv
VIDU_API_KEY
requires
env
bins
install
package
tsx
bins
[tsx]
package
ffmpeg-static
bins
[]

MedTravel Ad Maker

End-to-end pipeline that produces a polished 30-second medical-tourism ad video.

What it does

Given a destination (e.g. "Nanning, China"), a procedure (e.g. "dental implants"), and a brand name, this skill generates a complete 30-second ad with:

  1. Character continuity — One AI-generated protagonist appears in all 4 shots

(uses Vidu's reference2video so the same person shows up in every scene without per-shot drift).

  1. Cinematic visuals — 4 storyboarded shots:

- Pain point (high cost in patient's home country) - Modern destination clinic - Wellness recovery in scenic location - Triumphant outcome with brand CTA

  1. AI narration — Replicate Kokoro TTS (af_bella voice) generates

per-shot voiceover, time-aligned to each scene.

  1. Background music — Soft synthesized ambient pad (C-major triad,

low-pass filtered, fade in/out).

  1. Brand overlays — Top descriptive captions (so viewers understand the

story instantly) + bottom emerald-green brand text on each shot.

  1. Output — Final MP4 uploaded to your Cloudflare R2 bucket, plus all

intermediate clips for re-use.

How it works

Step 1: Wavespeed (Seedream 4.5) → 1 protagonist portrait → R2
Step 2: Vidu reference2video × 4 (parallel)  → 4 shot clips → R2
Step 3: Replicate Kokoro TTS × 4              → 4 narration clips
Step 4: ffmpeg concat                          → 30s silent video
Step 5: ffmpeg filter_complex                  → drawtext overlays + audio mix
Step 6: Upload final to R2

Cost & timing

Per run (one full 30s ad):

ItemCost
Wavespeed Seedream 4.5 (1 portrait)~$0.04
Vidu viduq2-pro reference2video × 4~$2.50 (250 credits)
Replicate Kokoro TTS × 4~$0.001
Total~$2.55

End-to-end runtime: ~3 minutes (most time is Vidu video generation in parallel).

Required environment variables

  • VIDU_API_KEY — Vidu Platform API key (https://platform.vidu.com)
  • WAVESPEED_API_KEY — Wavespeed.ai API key (for the protagonist image)
  • REPLICATE_API_KEY — Replicate token (for Kokoro TTS)
  • R2_ACCOUNT_ID, R2_ACCESS_KEY_ID, R2_SECRET_ACCESS_KEY,

R2_BUCKET_NAME, R2_PUBLIC_URL — Cloudflare R2 (S3-compatible) for storage

Required system binaries

  • node (≥ 18)
  • ffmpeg is bundled via the ffmpeg-static npm package — no system install needed.

Usage

# Customize the SHOTS array in make-xpilot-ad.ts with your storyboard,
# then run:
npx tsx make-xpilot-ad.ts

The script prints the final R2 URL at the end. To iterate on post-production (captions, narration, music) without re-spending Vidu credits, run:

npx tsx xpilot-ad-finalize.ts

This pulls the existing 4 video clips from R2, regenerates narration, and re-composites the final video. Free and fast (~45 seconds).

Example output

Final 30-second ad (8 MB MP4) — narration, ambient music, brand overlays: https://pub-22e3d3e3f43e400493bbd71306cae6bb.r2.dev/demo/medical-tourism-ad/v2/medtravel-final.mp4

Behind-the-scenes assets (all publicly hosted on R2):

  • Protagonist reference image (Wavespeed Seedream 4.5):

https://pub-22e3d3e3f43e400493bbd71306cae6bb.r2.dev/demo/medical-tourism-ad/v2/reference-protagonist.png

  • Shot 1 — Sticker shock:

https://pub-22e3d3e3f43e400493bbd71306cae6bb.r2.dev/demo/medical-tourism-ad/v2/shot-1-sticker-shock.mp4

  • Shot 2 — Nanning clinic:

https://pub-22e3d3e3f43e400493bbd71306cae6bb.r2.dev/demo/medical-tourism-ad/v2/shot-2-nanning-clinic.mp4

  • Shot 3 — Bama wellness:

https://pub-22e3d3e3f43e400493bbd71306cae6bb.r2.dev/demo/medical-tourism-ad/v2/shot-3-bama-wellness.mp4

  • Shot 4 — Detian triumph:

https://pub-22e3d3e3f43e400493bbd71306cae6bb.r2.dev/demo/medical-tourism-ad/v2/shot-4-detian-triumph.mp4

Notice the same protagonist appears in all 4 shots — that's the power of Vidu's reference2video mode, which this skill encapsulates.

Customization

To make this skill work for a different brand/vertical (e.g., "Mexican dental tourism", "Thai cosmetic surgery", "Korean LASIK"), edit:

  • REFERENCE_PROMPT — describe your protagonist
  • SHOTS[*].prompt — describe each scene
  • SHOTS[*].narration — what the voiceover says
  • SHOTS[*].brandText — bottom brand caption
  • SHOTS[*].topCaption — top descriptive caption

The pipeline (parallel submission, polling, R2 mirroring, ffmpeg composition) stays the same.

Why Reference-to-Video?

Vidu has three video generation modes:

ModeProsCons
text2videoSimpleEach shot's character looks different
img2videoVisual continuityHard to change scenes (just continues motion)
reference2videoSame character across scenesSlightly more setup

For multi-shot ads with a recurring protagonist, reference2video is the only mode that works. This skill encapsulates that workflow.

Known gotchas (saved you the debugging time)

  1. Vidu CloudFront URLs contain unencoded ; — don't URL-encode it,

that breaks the signature. Mirror to R2 immediately.

  1. OpenAI / OpenRouter quotas run out fast — this skill uses Replicate

Kokoro instead, which is dirt cheap.

  1. Replicate rate-limits accounts under $5 credit to 6 req/min — script

adds 11s delays between TTS calls.

  1. ffmpeg drawtext apostrophe escaping is unreliable — use full words

instead ("should not" instead of "shouldn't").

  1. ffmpeg drawtext % is parsed as variable — escape or use words ("60 percent").
  2. Multiple drawtext filters with commas in text break with , separator —

use ; + intermediate labels instead.

适合场景

01

调用多模型

02

代码和文本生成

03

Agent 推理流程

04

OpenRouter 模型接入

能力概览

能力 1

统一调用多种 LLM

能力 2

支持 Claude、Gemini、Kimi 等模型

能力 3

适合聊天、代码和推理任务

能力 4

可作为 Agent 模型调用入口

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

77.89%
按下载量换算642

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

可疑

权限和风险

敏感数据

该 Skill 可能接触密钥、Token、环境变量或敏感配置,应进入高风险复核队列,默认不自动发布。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。来源安全扫描存在 warning/failed 结果,不能写成本站确认安全。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills