Skip to content

Latest commit

 

History

History
88 lines (67 loc) · 3.69 KB

File metadata and controls

88 lines (67 loc) · 3.69 KB
domain archive
title Comfyui
verification metadata-normalized
source skill-pipeline
status draft
created 2026-05-09

背景

(此 lesson 从 skill comfyui 自动提取,待补全)

根因

(待补充)

修复

ComfyUI

Generate images, video, audio, and 3D content through ComfyUI using the official comfy-cli for setup/lifecycle and direct REST/WebSocket API for workflow execution.

What's in this skill

Reference docs (references/):

  • official-cli.md — every comfy ... command, with flags
  • rest-api.md — REST + WebSocket endpoints (local + cloud), payload schemas
  • workflow-format.md — API-format JSON, common node types, param mapping

Scripts (scripts/):

Script Purpose
_common.py Shared HTTP, cloud routing, node catalogs (don't run directly)
hardware_check.py Probe GPU/VRAM/disk → recommend local vs Comfy Cloud
comfyui_setup.sh Hardware check + comfy-cli + ComfyUI install + launch + verify
extract_schema.py Read a workflow → list controllable params + model deps
check_deps.py Check workflow against running server → list missing nodes/models
auto_fix_deps.py Run check_deps then comfy node install / comfy model download
run_workflow.py Inject params, submit, monitor, download outputs (HTTP or WS)
run_batch.py Submit a workflow N times with sweeps, parallel up to your tier
ws_monitor.py Real-time WebSocket viewer for executing jobs (live progress)
health_check.py Verification checklist runner — comfy-cli + server + models + smoke test
fetch_logs.py Pull traceback / status messages for a given prompt_id

Example workflows (workflows/): SD 1.5, SDXL, Flux Dev, SDXL img2img, SDXL inpaint, ESRGAN upscale, AnimateDiff video, Wan T2V. See workflows/README.md.

When to Use

  • User asks to generate images with Stable Diffusion, SDXL, Flux, SD3, etc.
  • User wants to run a specific ComfyUI workflow file
  • User wants to chain generative steps (txt2img → upscale → face restore)
  • User needs ControlNet, inpainting, img2img, or other advanced pipelines
  • User asks to manage ComfyUI queue, check models, or install custom nodes
  • User wants video/audio/3D generation via AnimateDiff, Hunyuan, Wan, AudioCraft, etc.

Architecture: Two Layers

┌─────────────────────────────────────────────────────┐
│ Layer 1: comfy-cli (official lifecycle tool)        │
│   Setup, server lifecycle, custom nodes, models     │
│   → comfy install / launch / stop / node / model    │
└─────────────────────────┬───────────────────────────┘
                          │
┌─────────────────────────▼───────────────────────────┐
│ Layer 2: REST/WebSocket API + skill scripts         │
│   Workflow execution, param injection, monitoring   │
│   POST /api/prompt, GET /api/view, WS /ws           │
│   → run_workflow.py, run_batch.py, ws_monitor.py    │
└─────────────────────────────────────────────────────┘

Why two layers? The official CLI is excellent for installation and server management but has minimal workflow execution support. The REST/WS API fills that gap — the scripts handle param injection, execution monitoring, and output

验证

(待补充)