Token导航 LogoToken导航TokenDH.com
研究检索执行命令clawhub未标认证来源可访问clear审计提醒

minimax-multimodal极小极大多模态

Agent Skill

minimax-multimodal 用于查找、检索和筛选相关信息,适合在 OpenClaw 中需要根据关键词、任务场景或来源线索快速定位候选结果时使用。可结合来源仓库、安装命令和原始 README 继续核验具体用法。安装前建议确认权限范围、维护状态,以及是否会触发联网、命令执行或文件读写。

总安装

121,181

周安装

4,855

GitHub Stars

17

下载量

39,228
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:minimax-multimodal(极小极大多模态)
来源仓库:https://github.com/minimax-ai-dev/minimax-multimodal
安装命令:
openclaw skills install minimax-multimodal
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install minimax-multimodal

简介

通过 MiniMax AI 平台调用 mmx 接口生成多种媒体内容。

  • 适合希望创建文本、图像、视频、语音或音乐的综合应用场景。
  • 用户可通过自然语言指令触发各类媒体生成任务。
  • 需确保拥有有效的 MiniMax 账户及相应 Token Plan 订阅。
  • 安装时请留意是否需要本地 Python 环境或 Node.js 运行时支持。

SKILL.md

name
mmx-cli
description
Use mmx to generate text, images, video, speech, and music via the MiniMax AI platform. Use when the user wants to create media content, chat with MiniMax models, perform web search, or manage MiniMax API resources from the terminal.

MiniMax CLI — Agent Skill Guide

Use mmx to generate text, images, video, speech, music, and perform web search via the MiniMax AI platform.

Prerequisites

# Install
npm install -g mmx-cli

# Auth (persisted to ~/.mmx/credentials.json)
mmx auth login --api-key sk-xxxxx

# Or pass per-call
mmx text chat --api-key sk-xxxxx --message "Hello"

Region is auto-detected. Override with --region global or --region cn.


Agent Flags

Always use these flags in non-interactive (agent/CI) contexts:

FlagPurpose
--non-interactiveFail fast on missing args instead of prompting
--quietSuppress spinners/progress; stdout is pure data
--output jsonMachine-readable JSON output
--asyncReturn task ID immediately (video generation)
--dry-runPreview the API request without executing
--yesSkip confirmation prompts

Commands

text chat

Chat completion. Default model: MiniMax-M2.7.

mmx text chat --message <text> [flags]
FlagTypeDescription
--message <text>string, required, repeatableMessage text. Prefix with role: to set role (e.g. "system:You are helpful", "user:Hello")
--messages-file <path>stringJSON file with messages array. Use - for stdin
--system <text>stringSystem prompt
--model <model>stringModel ID (default: MiniMax-M2.7)
--max-tokens <n>numberMax tokens (default: 4096)
--temperature <n>numberSampling temperature (0.0, 1.0]
--top-p <n>numberNucleus sampling threshold
--streambooleanStream tokens (default: on in TTY)
--tool <json-or-path>string, repeatableTool definition JSON or file path
# Single message
mmx text chat --message "user:What is MiniMax?" --output json --quiet

# Multi-turn
mmx text chat \
  --system "You are a coding assistant." \
  --message "user:Write fizzbuzz in Python" \
  --output json

# From file
cat conversation.json | mmx text chat --messages-file - --output json

stdout: response text (text mode) or full response object (json mode).


image generate

Generate images. Model: image-01.

mmx image generate --prompt <text> [flags]
FlagTypeDescription
--prompt <text>string, requiredImage description
--aspect-ratio <ratio>stringe.g. 16:9, 1:1
--n <count>numberNumber of images (default: 1)
--subject-ref <params>stringSubject reference: type=character,image=path-or-url
--out-dir <dir>stringDownload images to directory
--out-prefix <prefix>stringFilename prefix (default: image)
mmx image generate --prompt "A cat in a spacesuit" --output json --quiet
# stdout: image URLs (one per line in quiet mode)

mmx image generate --prompt "Logo" --n 3 --out-dir ./gen/ --quiet
# stdout: saved file paths (one per line)

video generate

Generate video. Default model: MiniMax-Hailuo-2.3. This is an async task — by default it polls until completion.

mmx video generate --prompt <text> [flags]
FlagTypeDescription
--prompt <text>string, requiredVideo description
--model <model>stringMiniMax-Hailuo-2.3 (default) or MiniMax-Hailuo-2.3-Fast
--first-frame <path-or-url>stringFirst frame image
--callback-url <url>stringWebhook URL for completion
--download <path>stringSave video to specific file
--asyncbooleanReturn task ID immediately
--no-waitbooleanSame as --async
--poll-interval <seconds>numberPolling interval (default: 5)
# Non-blocking: get task ID
mmx video generate --prompt "A robot." --async --quiet
# stdout: {"taskId":"..."}

# Blocking: wait and get file path
mmx video generate --prompt "Ocean waves." --download ocean.mp4 --quiet
# stdout: ocean.mp4

video task get

Query status of a video generation task.

mmx video task get --task-id <id> [--output json]

video download

Download a completed video by task ID.

mmx video download --file-id <id> [--out <path>]

speech synthesize

Text-to-speech. Default model: speech-2.8-hd. Max 10k chars.

mmx speech synthesize --text <text> [flags]
FlagTypeDescription
--text <text>stringText to synthesize
--text-file <path>stringRead text from file. Use - for stdin
--model <model>stringspeech-2.8-hd (default), speech-2.6, speech-02
--voice <id>stringVoice ID (default: English_expressive_narrator)
--speed <n>numberSpeed multiplier
--volume <n>numberVolume level
--pitch <n>numberPitch adjustment
--format <fmt>stringAudio format (default: mp3)
--sample-rate <hz>numberSample rate (default: 32000)
--bitrate <bps>numberBitrate (default: 128000)
--channels <n>numberAudio channels (default: 1)
--language <code>stringLanguage boost
--subtitlesbooleanInclude subtitle timing data
--pronunciation <from/to>string, repeatableCustom pronunciation
--sound-effect <effect>stringAdd sound effect
--out <path>stringSave audio to file
--streambooleanStream raw audio to stdout
mmx speech synthesize --text "Hello world" --out hello.mp3 --quiet
# stdout: hello.mp3

echo "Breaking news." | mmx speech synthesize --text-file - --out news.mp3

music generate

Generate music. Model: music-2.5. Responds well to rich, structured descriptions.

mmx music generate --prompt <text> [--lyrics <text>] [flags]
FlagTypeDescription
--prompt <text>stringMusic style description (can be detailed)
--lyrics <text>stringSong lyrics with structure tags. Use "\无\歌\词" for instrumental. Cannot be used with --instrumental
--lyrics-file <path>stringRead lyrics from file. Use - for stdin
--vocals <text>stringVocal style, e.g. "warm male baritone", "bright female soprano", "duet with harmonies"
--genre <text>stringMusic genre, e.g. folk, pop, jazz
--mood <text>stringMood or emotion, e.g. warm, melancholic, uplifting
--instruments <text>stringInstruments to feature, e.g. "acoustic guitar, piano"
--tempo <text>stringTempo description, e.g. fast, slow, moderate
--bpm <number>numberExact tempo in beats per minute
--key <text>stringMusical key, e.g. C major, A minor, G sharp
--avoid <text>stringElements to avoid in the generated music
--use-case <text>stringUse case context, e.g. "background music for video", "theme song"
--structure <text>stringSong structure, e.g. "verse-chorus-verse-bridge-chorus"
--references <text>stringReference tracks or artists, e.g. "similar to Ed Sheeran"
--extra <text>stringAdditional fine-grained requirements
--instrumentalbooleanGenerate instrumental music (no vocals). Cannot be used with --lyrics or --lyrics-file
--aigc-watermarkbooleanEmbed AI-generated content watermark
--format <fmt>stringAudio format (default: mp3)
--sample-rate <hz>numberSample rate (default: 44100)
--bitrate <bps>numberBitrate (default: 256000)
--out <path>stringSave audio to file
--streambooleanStream raw audio to stdout

At least one of --prompt or --lyrics is required.

# Simple usage
mmx music generate --prompt "Upbeat pop" --lyrics "La la la..." --out song.mp3 --quiet

# Detailed prompt with vocal characteristics
mmx music generate --prompt "Warm morning folk" \
  --vocals "male and female duet, harmonies in chorus" \
  --instruments "acoustic guitar, piano" \
  --bpm 95 \
  --lyrics-file song.txt \
  --out duet.mp3

# Instrumental (use --instrumental flag)
mmx music generate --prompt "Cinematic orchestral, building tension" --instrumental --out bgm.mp3

vision describe

Image understanding via VLM. Provide either --image or --file-id, not both.

mmx vision describe (--image <path-or-url> | --file-id <id>) [flags]
FlagTypeDescription
--image <path-or-url>stringLocal path or URL (auto base64-encoded)
--file-id <id>stringPre-uploaded file ID (skips base64)
--prompt <text>stringQuestion about the image (default: "Describe the image.")
mmx vision describe --image photo.jpg --prompt "What breed?" --output json

stdout: description text (text mode) or full response (json mode).


search query

Web search via MiniMax.

mmx search query --q <query>
FlagTypeDescription
--q <query>string, requiredSearch query
mmx search query --q "MiniMax AI" --output json --quiet

quota show

Display Token Plan usage and remaining quotas.

mmx quota show [--output json]

Tool Schema Export

Export all commands as Anthropic/OpenAI-compatible JSON tool schemas:

# All tool-worthy commands (excludes auth/config/update)
mmx config export-schema

# Single command
mmx config export-schema --command "video generate"

Use this to dynamically register mmx commands as tools in your agent framework.


Exit Codes

CodeMeaning
0Success
1General error
2Usage error (bad flags, missing args)
3Authentication error
4Quota exceeded
5Timeout
10Content filter triggered

Piping Patterns

# stdout is always clean data — safe to pipe
mmx text chat --message "Hi" --output json | jq '.content'

# stderr has progress/spinners — discard if needed
mmx video generate --prompt "Waves" 2>/dev/null

# Chain: generate image → describe it
URL=$(mmx image generate --prompt "A sunset" --quiet)
mmx vision describe --image "$URL" --quiet

# Async video workflow
TASK=$(mmx video generate --prompt "A robot" --async --quiet | jq -r '.taskId')
mmx video task get --task-id "$TASK" --output json
mmx video download --task-id "$TASK" --out robot.mp4

Configuration Precedence

CLI flags → environment variables → ~/.mmx/config.json → defaults.

# Persistent config
mmx config set --key region --value cn
mmx config show

# Environment
export MINIMAX_API_KEY=sk-xxxxx
export MINIMAX_REGION=cn

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

84.25%
按下载量换算33,050

安全审计

VirusTotal

通过

ClawScan

可疑

Static analysis

通过

权限和风险

执行命令

安装流程涉及命令执行,可能通过 openclaw skills install minimax-multimodal 联网下载 Skill 或依赖。用户安装前应确认命令来源、仓库内容和执行环境。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。来源安全扫描存在 warning/failed 结果,不能写成本站确认安全。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills