Token导航 LogoToken导航TokenDH.com
图像处理敏感数据clawhub未标认证来源可访问clear审计通过

ernie-image-visual-promptsmith厄尼 图像 视觉提示史密斯

Agent Skill

用于辅助界面设计、视觉规范、排版、配色、布局和交互体验优化。它适合让 Agent 根据产品场景整理页面结构、生成 UI 方案、检查视觉一致性或改进组件层级。使用时需要结合现有品牌、设计系统和用户任务,不应只堆装饰元素;涉及真实页面改动时,应通过截图或浏览器预览检查文本溢出、对齐和响应式表现。

总安装

2,060

周安装

85

GitHub Stars

1

下载量

673
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:ernie-image-visual-promptsmith(厄尼 图像 视觉提示史密斯)
来源仓库:https://github.com/yoimiya66/ernie-image-visual-promptsmith
安装命令:
openclaw skills install ernie-image-visual-promptsmith
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install ernie-image-visual-promptsmith

简介

ernie-image-visual-promptsmith 为海报、漫画与 UI 风格生成 ERNIE-Image 提示词。

  • 适合在 OpenClaw 中设计营销视觉、界面原型或数据可视化素材时使用。
  • 通过 clawhub 安装,结合百度 AI Studio 接口,支持风格迁移与元素组合优化。
  • 输出提示词需适配实际场景,避免过度修饰导致生成失败或失真。
  • 适用宿主包括 OpenClaw,接入前应确认版本、权限和运行环境要求。

SKILL.md

name
ernie-image-visual-promptsmith
description
Generate ERNIE-Image-Turbo images through Baidu AI Studio and craft ERNIE-Image prompts for posters, comics, infographics, ecommerce images, UI-style visuals, bilingual text rendering, structured layouts, negative prompts, generation settings, and use_pe decisions. Requires a user-provided AI Studio API key and is not an official Baidu skill.
metadata
openclaw
emoji
\F3A8
skillKey
ernie-image-visual-promptsmith
homepage
https://aistudio.baidu.com/account/accessToken
requires
env
anyBins
primaryEnv
BAIDU_AISTUDIO_API_KEY

ERNIE-Image Visual Promptsmith

Use this community skill to craft ERNIE-Image prompts and generate images through the AI Studio ERNIE-Image-Turbo endpoint. It is not official Baidu or ERNIE-Image software.

Decide the Mode

  • Generate immediately when the user asks to generate, draw, create, make an image, or uses equivalent Chinese generation wording.
  • Return prompt-only guidance when the user asks to optimize, rewrite, improve, or review a prompt.
  • Ask one concise question only if an exact visible text string, language, or required aspect ratio is missing and guessing would likely break the result.

API Endpoint

  • Base: https://aistudio.baidu.com/llm/lmapi/v3
  • Submit: POST /images/generations
  • Full URL: https://aistudio.baidu.com/llm/lmapi/v3/images/generations
  • Auth header: Authorization: bearer <BAIDU_AISTUDIO_API_KEY>
  • Platform header: X-Client-Platform: aistudio

API Key

  • Required environment variable: BAIDU_AISTUDIO_API_KEY
  • Get a key: https://aistudio.baidu.com/account/accessToken
  • If the key is missing, do not call the API. Tell the user to set BAIDU_AISTUDIO_API_KEY.

Triggers

  • Chinese examples: ERNIE image: <prompt>, Wenxin image: <prompt>, generate image: <prompt>, or equivalent Chinese wording for image generation.
  • English examples: ernie image: <prompt>, generate image: <prompt>, create image: <prompt>.
  • Treat text after the colon as the raw user prompt, improve it, choose a preset, then generate.
  • If the user asks to optimize, rewrite, improve, or review a prompt, return prompt-only guidance and do not call the API.

Prompt Workflow

  1. Classify the image style: photorealistic, anime/manga, text-in-image, concept art, abstract/artistic, layout/composition, poster, ecommerce, infographic, comic/storyboard, UI screenshot style, or character-consistent visual.
  2. Preserve immutable constraints: exact in-image text, language, subject count, character identity, spatial relationships, size, style, and forbidden elements.
  3. Build the core prompt in five parts: subject -> action/context -> style -> lighting -> quality.
  4. For layout-sensitive requests, append composition -> exact text -> spatial placement.
  5. Keep in-image writing short when possible. Turn paragraphs into titles, labels, badges, or numbered lines.
  6. For text rendering, put exact wording in quotes and specify placement, font weight, alignment, color, background contrast, and whitespace.
  7. Choose a preset from auto, text-poster, infographic, comic, product, ui, photo, concept, or abstract.
  8. Before generation, state:
Final Prompt: <prompt>
Preset: <preset>
use_pe: <true or false>
Size: <size>
Reason: <why these settings fit ERNIE-Image>

Generation Workflow

Use the bundled Python script. Prefer python3; on Windows use python or py if needed.

python3 {baseDir}/scripts/generate.py --prompt "<FINAL_PROMPT>" --preset <preset>

For exact text, bilingual labels, UI, flowcharts, signs, comics, or already detailed prompts, pass --no-use-pe.

python3 {baseDir}/scripts/generate.py --prompt "<FINAL_PROMPT>" --preset text-poster --no-use-pe

The script prints IMAGE_URL:<url> for URL responses and MEDIA:<absolute_path> for each saved image. Return the saved media path to the user.

If BAIDU_AISTUDIO_API_KEY is missing, tell the user to get a key from https://aistudio.baidu.com/account/accessToken and set BAIDU_AISTUDIO_API_KEY.

Submit Payload

{
  "model": "ERNIE-Image-Turbo",
  "prompt": "<FINAL_PROMPT>",
  "n": 1,
  "response_format": "url",
  "size": "1024x1024",
  "seed": 42,
  "use_pe": true,
  "num_inference_steps": 8,
  "guidance_scale": 1.0
}

Download and Output

  • response_format=url returns image URLs in data[]; the script prints IMAGE_URL:<url>.
  • The script downloads each URL immediately and saves the image locally.
  • The script prints MEDIA:<absolute_path> for OpenClaw/ClawHub auto-attach.
  • URLs may expire; the local file remains available after download.
  • Output names are generated as ernie-image-<timestamp>-<index>.<ext>.
  • Do not pass user-controlled filenames to shell commands.

Defaults

  • Model: ERNIE-Image-Turbo
  • Preset: auto
  • Count: 1
  • Response format: url
  • Seed: 42
  • text-poster, infographic, comic, product, and ui presets default to use_pe=false.
  • photo, concept, and abstract presets default to use_pe=true.

Negative Prompt Rules

  • Do not add text, letters, typography, Chinese text, or English text when the user wants readable writing.
  • Prefer precise negatives: distorted text, misspelled words, duplicated letters, unreadable typography, warped layout, cropped title, low contrast, blurry details, inconsistent panels, artifacts.
  • The API does not expose a separate negative prompt field in this skill. Express exclusions as natural language constraints inside the prompt, such as "avoid cluttered background" or "no visible watermark".

Retry Strategy

  • Text errors: reduce the amount of visible text, quote exact words once, add stronger placement and contrast, then use --no-use-pe.
  • Layout errors: simplify object count, name each region, use grid/split-screen/foreground/background terms, then keep the same seed.
  • Weak style: add camera/lens, art movement, medium, color temperature, material texture, and lighting direction.
  • Cluttered image: remove secondary elements, add negative space, use "avoid cluttered background", and switch to a simpler preset if needed.

References

  • Read references/api.md for parameters, command examples, and endpoint mapping.
  • Read references/prompt-architecture.md for ERNIE-Image prompt templates.
  • Read references/examples.md for acceptance-style examples.

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

78.13%
按下载量换算526

安全审计

VirusTotal

未展示

ClawScan

通过

Static analysis

通过

权限和风险

敏感数据

该 Skill 可能接触密钥、Token、环境变量或敏感配置,应进入高风险复核队列,默认不自动发布。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills