Anima, Krea-2, Qwen-Image 2.1 & YuE2 INT8 Workflows Collection
This repository provides optimized ComfyUI workflows for the Anima, Krea-2, Qwen-Image 2.1, and YuE2 model families.
The image-generation workflows utilize INT8 quantized variants for efficient Text-to-Image generation with a low VRAM footprint. Anima provides advanced structural control and image editing capabilities, while Qwen-Image 2.1 provides image generation and editing with mask-based local editing, reference images, and native RGBA support.
The repository also includes a YuE2 music-generation workflow for generating complete songs from style prompts and lyrics.
Workflow Files & Overview
The repository contains the following production-ready workflow files.
(Note: YYMM.x represents the release year, month, and version number, e.g., 2606.1 or 2607.1.)
| # | File Name | Workflow Target | Key Features & Quick Notes |
|---|---|---|---|
| 1 | Anima_int8-javanoYYMM.x.json | Anima Text-to-Image (INT8) | INT8 image generation with VNCCS, LLLite-based Pose/Canny/Depth control, SAM 3.1 automatic masking, and inpainting. |
| 2 | Krea2_int8_t2i-javanoYYMM.x.json | Krea-2 Text-to-Image (INT8) | Highly optimized INT8 workflow for Krea-2, designed for fast and efficient image generation. |
| 3 | QwenImage2.1_int8-javanoYYMM.x.json | Qwen-Image 2.1 Text-to-Image / Image Editing (INT8) | INT8 workflow with sampling preview, Face Detailer, mask-based editing, reference image editing, and Ollama-optimized prompt rewriting based on Alibaba's official Qwen-Image prompt skills. |
| 4 | YuE2-javanoYYMM.x.json | YuE2 Music Generation | Generates complete songs from style prompts and lyrics using the YuE2-3B model, with vocals and instrumental accompaniment. |
Detailed Workflow Breakdown
1. Anima Text-to-Image / Control / Inpainting (Anima INT8)
Designed for high-quality static image generation using the INT8 version of Anima, with additional structural control and automated image editing.
- INT8 Inference: Reduced VRAM usage and improved inference efficiency.
- Standard Text-to-Image: Generate images directly from text prompts using Anima.
- VNCCS Control: Provides Pose Doll-based character and pose control.
- LLLite Control: Supports structural guidance using Pose, Canny, and Depth conditioning.
- SAM 3.1 Masking: Automatically generates masks for selected subjects or regions.
- Inpainting: Combines SAM 3.1 automatic masks with Anima LLLite conditioning for targeted image regeneration.
Anima INT8:
https://huggingface.co/Bedovyy/Anima-INT8
Anima LLLite:
https://huggingface.co/Comfy-Org/Anima-LLLite
VNCCS:
https://github.com/AHEKOT/ComfyUI_VNCCS
SAM 3.1:
https://huggingface.co/Comfy-Org/sam3.1
Supported control workflows include:
- Standard Text-to-Image
- Pose Doll
- Canny
- Depth
- Automatic subject masking
- Region-based inpainting
2. Krea-2 Text-to-Image (Krea-2 INT8)
A dedicated INT8 workflow designed to utilize the aesthetic capabilities of Krea-2 within an efficient quantized framework.
- INT8 Acceleration:
Krea2_int8_t2i-javanoYYMM.x.jsonleverages INT8 quantization for efficient inference. - Cinematic & Spatial Optimization: Configured to maximize Krea-2's strengths in lighting, realistic rendering, and composition.
- Efficient Inference: Quantization reduces memory requirements while maintaining high image quality.
3. Qwen-Image 2.1 Text-to-Image / Image Editing (Qwen-Image 2.1 INT8)
A dedicated INT8 ComfyUI workflow for Qwen-Image 2.1, optimized for efficient image generation and image editing.
- INT8 Inference: Uses the INT8 version of Qwen-Image 2.1 for reduced VRAM usage and efficient inference.
- Sampling Preview: Provides an intermediate sampling preview for monitoring the generation process.
- Face Detailer: Adds dedicated face-detail refinement after generation.
- Mask Editing: Enables localized image editing by specifying the target region with a mask.
- Image Editing: Supports Qwen-Image 2.1 image-editing capabilities.
- Reference Images: Supports reference-based image generation and editing.
- Native RGBA: Supports native RGBA image generation and transparent backgrounds.
- Ollama Prompt Optimization: Uses an Ollama-based prompt rewriting workflow optimized around Alibaba's official Qwen-Image prompt-rewriting skills.
- Official Prompt Skills: The prompt optimization follows the structure and principles of Qwen-Image 2.1's official prompt-rewriting system prompts for Text-to-Image and Image-to-Image editing.
Qwen-Image 2.1:
https://huggingface.co/Comfy-Org/Qwen-Image-2.1
Qwen-Image 2.1 Project:
https://github.com/QwenLM/Qwen-Image-2.1
Supported editing workflows include:
- Text-to-Image
- Image-to-Image
- Mask-based local editing
- Face Detailer
- Reference image editing
- Native RGBA / transparent image generation
4. YuE2 Music Generation (YuE2)
A dedicated ComfyUI workflow for YuE2-3B, an open music-generation model from the Multimodal Art Projection (m-a-p) team.
YuE2 generates complete songs from a style prompt and lyrics, producing vocals and instrumental accompaniment.
- Text-to-Music: Generate complete songs from style and lyrics.
- Vocals & Instrumentation: Produces full songs with vocals and accompaniment.
- Symbolic Planning: Uses symbolic musical planning for melody and harmony.
- Complete Song Generation: Designed for full-song generation rather than short audio clips.
- High-Quality Audio: Supports high-quality stereo music output.
YuE2 Model:
https://huggingface.co/Comfy-Org/YuE2
YuE2 Project:
https://github.com/multimodal-art-projection/YuE
Typical input structure:
Style:
J-pop, female vocal, energetic electronic rock, bright synths, live drums
Lyrics:
[Verse]
...
[Chorus]
...