Token导航 LogoToken导航TokenDH.com
待分类权限需确认github未标认证来源可访问许可证需确认审计未展示

comfyui-character-gencomfyui 角色生成

Agent Skill

comfyui-character-gen 用于处理 GitHub 仓库、Issue、Pull Request 和代码协作信息,适合在 Codex、Claude、Cursor、Gemini CLI 中需要围绕仓库状态、代码变更或协作事项进行整理时使用。可结合来源仓库、安装命令和原始 README 继续核验具体用法。安装前建议确认权限范围、维护状态,以及是否会触发联网、命令执行或文件读写。

总安装

2,064

周安装

86

GitHub Stars

50

下载量

688
CodexClaudeCursorGemini CLI

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

unknown

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:comfyui-character-gen(comfyui 角色生成)
来源仓库:https://github.com/mckruz/comfyui-expert
仓库路径:skills/comfyui-character-gen
安装命令:
npx skills add https://github.com/mckruz/comfyui-expert --skill comfyui-character-gen
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 npx skills 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

skills.shnpx skills
npx skills add https://github.com/mckruz/comfyui-expert --skill comfyui-character-gen

简介

该技能暂无可用文档,仅提供 GitHub 链接供查看详情。

  • 可能涉及角色 LoRA 训练或图像生成相关功能。
  • 建议访问源码仓库查看最新说明和使用示例。
  • 当前无法提供具体功能描述或操作指南。适用宿主包括 Codex、Claude、Cursor、Gemini CLI,接入前应确认版本、权限和运行环境要求。
  • comfyui-character-gen 属于待分类类 Skill,可作为该场景下的辅助能力补充。

SKILL.md

ComfyUI Character Generation Expert

Build production-ready ComfyUI workflows for consistent character generation across image, video, and voice modalities.

Quick Decision: Which Approach?

Starting from reference images (like 3D renders)?InfiniteYou (state-of-the-art 2025) or InstantID + IP-Adapter (proven, lower VRAM)

Need highest identity fidelity?FLUX.2 (NEW 2026: up to 10 ref images) or PuLID Flux II (no model pollution)

Want iterative editing without retraining?FLUX Kontext (context-aware, maintains consistency across edits)

Creating video content?LTX-2 (NEW 2026: 4K production-ready), Wan 2.2 MoE (film-level), or FramePack (60-sec on 6GB!)

Need voice for character?TTS Audio Suite (unified platform, 23 languages) or F5-TTS Cross-Lingual (NEW 2026)

Core Workflow Patterns

Pattern 1: Zero-Shot Character Generation (No Training)

Best for: Quick iteration, 3D-to-photorealism conversion, limited reference images

Load Reference Face → InstantID + IP-Adapter FaceID → ControlNet Pose → KSampler → FaceDetailer → Upscale

Critical settings:

  • CFG: 4-5 (prevents burning with InstantID)
  • Resolution: 1016×1016 (avoids watermark artifacts)
  • IP-Adapter weight: 0.6-0.8
  • InstantID noise injection: 35% to negative

See references/workflows.md for complete node configurations.

Pattern 2: LoRA + Identity Methods (Maximum Consistency)

Best for: Production work, character series, video generation base

Train LoRA → Load LoRA + Checkpoint → Add InstantID/PuLID → Generate → FaceDetailer → ReActor (optional) → Upscale

Training requirements:

  • 15-30 images, varied poses/expressions/lighting
  • Unique trigger word (e.g., "sage_character")
  • See references/lora-training.md for full parameters

Pattern 3: Video Generation Pipeline

Best for: Talking heads, character animation, promotional content

Generate/Load Hero Image → Wan 2.1 I2V OR AnimateDiff → FaceDetailer per frame → Frame Interpolation → Video Combine

Model selection:

  • Wan 2.1 14B: Best quality, 24GB+ VRAM, slower
  • Wan 2.1 1.3B: 8GB VRAM, good quality, faster
  • AnimateDiff Lightning: Fastest, best for iteration

Pattern 4: Talking Head with Voice

Best for: Character dialogue, presentations, social content

Two approaches available:

Approach 1 (Image → Talking Head):
Character Portrait → Generate Audio → SadTalker/LivePortrait → CodeFormer Enhancement → Final Video

Approach 2 (Video → Add Voice):
Existing Video → Generate Audio → Wav2Lip Lip-Sync → CodeFormer Enhancement → Final Video

See references/talking-head-workflows.md for complete workflows and references/voice-synthesis.md for voice creation options.

Model Recommendations (2026 Updated)

Image Generation

Use CaseModelNotes
Best photorealismFLUX.1-devSlow but superior quality
Multi-reference consistencyFLUX.2NEW 2026: Up to 10 ref images, strong identity preservation
Fast iterationRealVisXL V5.0Good balance speed/quality
Character editingFLUX KontextContext-aware, maintains consistency across edits
Iterative refinementFLUX Kontext Pro/Max8x faster than GPT-Image (API)

Identity Preservation (2026 State-of-Art)

MethodBest ForVRAMNotes
FLUX.2Multi-reference consistency24GB+NEW 2026: Up to 10 ref images, branded content
InfiniteYouHighest identity match24GBICCV 2025 Highlight, SIM/AES variants
FLUX KontextIterative editing12-32GBBuilt-in consistency, no retraining
PuLID Flux IIDual characters, no pollution24-40GBContrastive alignment solves model pollution
AuraFaceCommercial identity encoding12GBNEW 2026: Open-source ArcFace alternative
InstantIDStyle transfer, 3D→realistic12GBMaintenance mode but still excellent
IP-Adapter FaceIDSpeed, lower VRAM6GB+Good baseline approach

Video Generation

ModelQualitySpeedVRAMNotes
LTX-2★★★★★Medium16GB+NEW 2026: First open-source 4K audio+video, production-ready
Wan 2.2 MoE★★★★★Slow24GB+Film-level aesthetics, first+last frame control
FramePack★★★★★Medium6GB60-sec videos, VRAM-invariant breakthrough
Wan 2.1 1.3B★★★★Medium8GB+Consumer-friendly
AnimateDiff V3★★★Fast8GBMotion/camera LoRAs, infinite length

Voice/TTS

ToolLicenseQualityFeatures
TTS Audio SuiteMulti★★★★★Unified platform, 23 languages, emotion control
F5-TTSMIT★★★★Zero-shot from <15 sec samples, Cross-Lingual 2026
ChatterboxMIT★★★★★Paralinguistic tags ([laugh], [sigh]), 4 voices
IndexTTS-2MIT★★★★8-emotion vector control
ElevenLabsCommercial★★★★★Production quality (API)

Essential Custom Nodes

Install via ComfyUI-Manager:

ComfyUI-Manager              # Must install first
ComfyUI_IPAdapter_plus       # IP-Adapter and FaceID
ComfyUI_InstantID            # InstantID workflow
ComfyUI-Impact-Pack          # FaceDetailer
ComfyUI-ReActor              # Face swapping
ComfyUI-AnimateDiff-Evolved  # Video generation
ComfyUI-VideoHelperSuite     # Video I/O
comfyui_controlnet_aux       # Pose/depth preprocessors
ComfyUI_UltimateSDUpscale    # Tiled upscaling
ComfyUI-Frame-Interpolation  # Smooth video

RTX 50 Series Optimization (NEW 2026)

With 32GB VRAM on RTX 5090, run most workflows without optimization. ComfyUI v0.8.1 adds major RTX 50 Series enhancements:

Launch flags: --highvram --fp8_e4m3fn-unet

NEW v0.8.1 Features:

  • NVFP4/NVFP8 precision formats: 3x faster performance, 60% VRAM reduction on RTX 50 Series
  • Weight streaming: Uses system RAM when VRAM exhausted, enables larger models on mid-range GPUs
  • Enable tiled VAE for 8K+ upscaling
  • Batch 4× 1024×1024 generations in parallel
  • Run Wan 2.2 14B + LTX-2 natively
  • Use FP8 quantization for FLUX (50% VRAM reduction)

Workflow Generation Process

When building a workflow for a user:

  1. Clarify the goal: Image only? Video? With voice? What's the source material?
  2. Select the pipeline pattern from above based on requirements
  3. Generate the workflow following node configurations in references/workflows.md
  4. Include model downloads with exact filenames and paths from references/models.md
  5. Provide parameter recommendations specific to their hardware/use case

Reference Files

  • references/research-log.md - Latest techniques: InfiniteYou, FLUX Kontext, PuLID Flux II, Wan 2.2 MoE, FramePack, FLUX.2, LTX-2.3, Wan 2.6, Qwen3-TTS, and more
  • references/models.md - Complete model list with HuggingFace/Civitai links, file paths, and compatibility notes
  • references/workflows.md - Detailed node-by-node workflow templates for each pattern
  • references/lora-training.md - LoRA training guide with Kohya/AI-Toolkit parameters
  • references/voice-synthesis.md - Voice cloning, TTS, and lip-sync pipeline details
  • references/talking-head-workflows.md - Complete talking head workflows: Image→Talking Head (SadTalker, LivePortrait) and Video→Add Voice (Wav2Lip) with production scripts
  • references/evolution.md - Update sources, changelog, and user-specific learnings

Skill Evolution

This skill is designed to evolve. When helping the user:

Before starting a workflow:

  • Check if new models have dropped that might be better (search HuggingFace/Civitai if uncertain)
  • Consider if user's past successes/failures inform the approach

After completing a workflow:

  • Note what worked well or poorly for future reference
  • If user discovers better settings, update the relevant reference file

Proactive updates:

  • When the user mentions a new model or technique, research and integrate it
  • Periodically suggest checking for updates to key dependencies

See references/evolution.md for monitoring sources and update protocols.

Example: 3D Render to Photorealistic Character

For converting stylized 3D renders (like game/VN characters) to photorealistic images:

Recommended approach: InstantID + IP-Adapter FaceID on FLUX

1. Load 3D render reference (best quality, front-facing)
2. Apply InstantID (extracts identity + facial keypoints)
3. Apply IP-Adapter FaceID Plus V2 (weight 0.7)
4. Use FLUX.1-dev checkpoint
5. Prompt: "photorealistic portrait, detailed skin texture, natural lighting, [character description]"
6. CFG: 4-5, Steps: 25-30
7. FaceDetailer pass (denoise 0.35)
8. Upscale with 4x-UltraSharp

This converts the stylized look to photorealism while preserving the core identity features.

适合场景

01

用户想查找某类 Agent Skill 时

02

需要根据任务场景推荐可安装能力包时

03

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

Codex

37.6%
按下载量换算259

Claude

30.02%
按下载量换算207

Cursor

19.48%
按下载量换算134

Gemini CLI

10.38%
按下载量换算71

安全审计

暂无安全审计结果可展示。

权限和风险

权限需确认

当前来源未能明确判断权限范围,默认进入异常复核队列。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills