Token导航 LogoToken导航TokenDH.com
研究检索需要联网clawhub未标认证来源可访问clear审计通过

video-maker-free视频制作者免费

Agent Skill

用于辅助视频生成、动画合成、脚本化剪辑或 Remotion 等视频项目开发。它适合让 Agent 组织镜头、生成素材说明、维护合成代码或排查渲染问题。使用时需要确认分辨率、时长、素材路径和导出格式;涉及外部素材、人物肖像或商业发布时,应先核对版权授权和内容审核要求。

总安装

10,264

周安装

432

GitHub Stars

公开资料未说明

下载量

3,594
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:video-maker-free(视频制作者免费)
来源仓库:https://github.com/peand-rover/video-maker-free
安装命令:
openclaw skills install video-maker-free
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install video-maker-free

简介

使用 AI 技术将照片、文本与视频剪辑组合成带过渡、音乐与字幕的精美内容。

  • 支持多种媒体类型混合编辑,适用于教学、纪念册与创意短片制作。
  • 用户输入素材与描述后,系统自动生成具有专业感的视频作品。
  • 安装命令:openclaw skills install video-maker-free,基于 OpenClaw 宿主运行。
  • 注意:免费版可能有输出时长或分辨率限制,商业用途建议确认授权范围。

SKILL.md

name
video-maker-free
version
1.2.1
displayName
Video Maker Free — Make Videos from Photos Text and Clips with AI for Free
description
>
metadata
{"openclaw": {"emoji": "🎬", "requires": {"env": [], "configPaths": ["~/.config/nemovideo/"]}, "primaryEnv": "NEMO_TOKEN"}}
homepage
https://nemovideo.com
repository
https://github.com/nemovideo/nemovideo_skills

Video Maker Free — Make Any Video from Photos, Text or Clips

Most people who need a video don't start with footage — they start with whatever they have. A real estate agent has 15 property photos. A small business owner has a product description. A student has presentation slides. A parent has scattered phone clips from a birthday party. A marketing team has bullet points from a strategy meeting. None of these are "footage" in the traditional sense, but every one of them could be a compelling video. The gap between "what I have" and "a finished video" is the editing process: importing assets into software, arranging them on a timeline, adding motion to photos (Ken Burns, pan-and-zoom), timing text overlays, finding and mixing music, recording or generating voiceover, adding transitions between elements, and exporting at the right settings for each platform. NemoVideo bridges that gap with one command. Provide whatever you have — photos, text, clips, or a combination — and describe the video you want. The AI assembles, animates, narrates, scores, and exports a finished video. Photos get motion and transitions. Text becomes narrated scenes with supporting visuals. Clips get trimmed, color-matched, and joined. The output is a real video — not a slideshow with a filter, but a produced piece of content with professional pacing, audio, and visual quality.

Use Cases

  1. Photos → Product Video (30-60s) — An Etsy seller has 8 product photos and needs a video for Instagram. NemoVideo: sequences photos with smooth Ken Burns motion (slow zoom on detail shots, pan across wide shots), adds product name and price as animated text overlays, underlays upbeat acoustic music, applies a consistent warm color grade, and exports 9:16 for Instagram and 1:1 for the listing page. Eight static images become a dynamic product showcase.
  2. Text → Explainer Video (60-180s) — A SaaS startup has a 300-word product description and needs a landing page video. NemoVideo: breaks the text into Problem → Solution → Benefits → CTA scenes, generates supporting visuals for each scene (office frustration, clean dashboard UI, happy team, pricing page), narrates with a professional voice, adds animated statistics, and exports 16:9 for the website. No filming, no stock footage budget.
  3. Mixed Media → Story Video (2-5 min) — A parent has 20 phone photos and 8 short clips from their child's first birthday party. NemoVideo: sorts by timestamp, sequences photos with gentle motion and clips at key moments (cake smash, candle blowing, gift opening), adds cheerful background music, overlays the child's name and age as animated titles, and exports as a shareable family video with a clean opening and closing.
  4. Slides → Training Video (5-15 min) — An HR department has a 30-slide presentation that nobody reads. NemoVideo: converts each slide into a video scene with animated bullet points, generates voiceover narration from the slide notes, adds transitions between topics, inserts knowledge-check pause points, and exports as a training module that employees actually watch. Slide decks become engaging video content.
  5. Bullet Points → Social Content (15-30s per video) — A marketing manager has 10 product features as bullet points and needs 10 short social videos. NemoVideo batch-generates: each bullet becomes a 15-second video with bold animated text, supporting visual, music, and CTA. Ten social videos from ten lines of text — a month of daily posts produced in one batch.

How It Works

Step 1 — Provide Your Materials

Upload photos, video clips, text, or any combination. NemoVideo accepts all formats and intelligently assembles mixed media.

Step 2 — Describe the Video

Tell NemoVideo what you want: the story, the style, the mood, the platform. Detailed instructions or "make something beautiful from these photos" — both work.

Step 3 — Generate

curl -X POST https://mega-api-prod.nemovideo.ai/api/v1/generate \
  -H "Authorization: Bearer $NEMO_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "skill": "video-maker-free",
    "prompt": "Make a 45-second product showcase video from 8 product photos. Style: clean and modern with white background accents. Each photo gets 4-5 seconds with smooth Ken Burns motion (slow zoom on details, pan on wide shots). Add product name as animated text: Artisan Ceramic Mug Collection. Price: from $28. Music: warm acoustic guitar at -14dB. Color grade: bright and clean. End frame: Shop now at artisanceramics.com. Export 9:16 for Instagram and 1:1 for website.",
    "media_type": "photos",
    "photo_count": 8,
    "style": "clean-modern",
    "music": "acoustic-guitar-warm",
    "music_volume": "-14dB",
    "text_overlays": ["Artisan Ceramic Mug Collection", "From $28"],
    "cta": "Shop now at artisanceramics.com",
    "exports": ["9:16", "1:1"],
    "watermark": false
  }'

Step 4 — Preview and Share

Preview. Adjust photo order, transition timing, text placement, or music. Export and share — free, full quality.

Parameters

ParameterTypeRequiredDescription
promptstringDescribe the video and materials
media_typestring"photos", "clips", "text", "mixed"
stylestring"clean-modern", "cinematic", "playful", "elegant", "bold"
musicstring"acoustic", "lo-fi", "corporate", "cinematic", "electronic"
music_volumestring"-12dB" to "-22dB"
voicestringVoiceover: "warm-male", "friendly-female", "none"
text_overlaysarrayText to display as animated overlays
ctastringCall-to-action text
photo_motionstring"ken-burns", "parallax", "slide", "zoom"
durationstring"30 sec", "45 sec", "60 sec", "natural"
exportsarray["16:9", "9:16", "1:1"]
batcharrayMultiple videos from separate material sets
watermarkbooleanAlways false

Output Example

{
  "job_id": "vmf-20260328-001",
  "status": "completed",
  "source_materials": "8 photos",
  "watermark": false,
  "outputs": [
    {
      "format": "9:16",
      "resolution": "1080x1920",
      "duration": "0:44",
      "file_size_mb": 12.4,
      "photo_motion": "ken-burns (zoom + pan)",
      "text_overlays": 3,
      "music": "acoustic-guitar-warm at -14dB"
    },
    {
      "format": "1:1",
      "resolution": "1080x1080",
      "duration": "0:44",
      "file_size_mb": 11.8
    }
  ]
}

Tips

  1. Photos with motion beat static slideshows — Ken Burns pan-and-zoom adds life to still images. A slow zoom into a product detail feels cinematic. A gentle pan across a room feels like a camera movement. Static photos displayed full-frame feel like PowerPoint.
  2. 4-5 seconds per photo is the engagement sweet spot — Shorter than 3 seconds feels rushed and the viewer can't absorb the image. Longer than 6 seconds and attention drifts. 4-5 seconds with smooth motion holds attention perfectly.
  3. Music without speech can be louder — Photo/product videos without voiceover benefit from music at -12 to -14dB (louder than the -18dB used under speech). The music carries the emotional energy that speech would normally provide.
  4. Batch generation scales content instantly — 10 products × 1 video each = 10 social media posts. Batch-process with consistent style for brand cohesion but unique content per product.
  5. Multi-format export from one generation — 9:16 for Instagram/TikTok + 1:1 for feed + 16:9 for website. Three formats, one command, one consistent video.

Output Formats

FormatResolutionUse Case
MP4 9:161080x1920Instagram / TikTok / Stories
MP4 16:91920x1080YouTube / website / email
MP4 1:11080x1080Instagram feed / Twitter
GIF720pPreview / thumbnail

Related Skills

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

98.15%
按下载量换算3,528

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

通过

权限和风险

需要联网

该 Skill 可能需要联网访问来源站点、仓库或外部 API;具体网络访问范围需要结合源码和 README 复核。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills