ComfyUI
Node-graph engine for diffusion image, video, audio and 3D models
Builds generation pipelines as a visual node graph and runs them locally for image (SD 1.5, SDXL, SD3.5, Flux.1 and Flux.2, Qwen Image), video (Wan 2.x, LTX-Video, HunyuanVideo), audio (ACE-Step, Stable Audio) and 3D (Hunyuan3D) models, with a local API and an App Mode that exposes a workflow as a simple UI. Runs on NVIDIA, AMD, Intel, Apple Silicon and Ascend. For professionals who want control over every parameter.
Strengths
- Asynchronous weight streaming runs large models on 4 GB VRAM plus 8 GB RAM
- Workflows saved as JSON and recoverable from generated media metadata
- Runs fully offline; --offline disables the paid API nodes
- Loads checkpoints, separate diffusion models, VAEs, text encoders, LoRAs, ControlNets
Weaknesses
- Commits outside stable tags can break many custom nodes; stable releases roughly biweekly
- GPL-3.0 license constrains embedding in proprietary products
- NVIDIA 20-series and newer require PyTorch built with CUDA 13.0 or above
- Paid partner and API nodes stay on unless --offline or --disable-partner-nodes is set
- RAM ≥ 8 GB
- GPU optional
- Models: Stable Diffusion 1.5, SDXL, SD3.5, Flux.1 and Flux.2, Qwen Image and Qwen Image Edit, Wan 2.1/2.2, LTX-Video 2