Anima, Krea-2, Qwen-Image 2.1 & YuE2 INT8 Workflows Collection

This repository provides optimized ComfyUI workflows for the Anima, Krea-2, Qwen-Image 2.1, and YuE2 model families.

The image-generation workflows utilize INT8 quantized variants for efficient Text-to-Image generation with a low VRAM footprint. Anima provides advanced structural control and image editing capabilities, while Qwen-Image 2.1 provides image generation and editing with mask-based local editing, reference images, and native RGBA support.

The repository also includes a YuE2 music-generation workflow for generating complete songs from style prompts and lyrics.


Workflow Files & Overview

The repository contains the following production-ready workflow files.

(Note: YYMM.x represents the release year, month, and version number, e.g., 2606.1 or 2607.1.)

# File Name Workflow Target Key Features & Quick Notes
1 Anima_int8-javanoYYMM.x.json Anima Text-to-Image (INT8) INT8 image generation with VNCCS, LLLite-based Pose/Canny/Depth control, SAM 3.1 automatic masking, and inpainting.
2 Krea2_int8_t2i-javanoYYMM.x.json Krea-2 Text-to-Image (INT8) Highly optimized INT8 workflow for Krea-2, designed for fast and efficient image generation.
3 QwenImage2.1_int8-javanoYYMM.x.json Qwen-Image 2.1 Text-to-Image / Image Editing (INT8) INT8 workflow with sampling preview, Face Detailer, mask-based editing, reference image editing, and Ollama-optimized prompt rewriting based on Alibaba's official Qwen-Image prompt skills.
4 YuE2-javanoYYMM.x.json YuE2 Music Generation Generates complete songs from style prompts and lyrics using the YuE2-3B model, with vocals and instrumental accompaniment.

Detailed Workflow Breakdown

1. Anima Text-to-Image / Control / Inpainting (Anima INT8)

Designed for high-quality static image generation using the INT8 version of Anima, with additional structural control and automated image editing.

  • INT8 Inference: Reduced VRAM usage and improved inference efficiency.
  • Standard Text-to-Image: Generate images directly from text prompts using Anima.
  • VNCCS Control: Provides Pose Doll-based character and pose control.
  • LLLite Control: Supports structural guidance using Pose, Canny, and Depth conditioning.
  • SAM 3.1 Masking: Automatically generates masks for selected subjects or regions.
  • Inpainting: Combines SAM 3.1 automatic masks with Anima LLLite conditioning for targeted image regeneration.

Anima INT8:
https://huggingface.co/Bedovyy/Anima-INT8

Anima LLLite:
https://huggingface.co/Comfy-Org/Anima-LLLite

VNCCS:
https://github.com/AHEKOT/ComfyUI_VNCCS

SAM 3.1:
https://huggingface.co/Comfy-Org/sam3.1

Supported control workflows include:

  • Standard Text-to-Image
  • Pose Doll
  • Canny
  • Depth
  • Automatic subject masking
  • Region-based inpainting

2. Krea-2 Text-to-Image (Krea-2 INT8)

A dedicated INT8 workflow designed to utilize the aesthetic capabilities of Krea-2 within an efficient quantized framework.

  • INT8 Acceleration: Krea2_int8_t2i-javanoYYMM.x.json leverages INT8 quantization for efficient inference.
  • Cinematic & Spatial Optimization: Configured to maximize Krea-2's strengths in lighting, realistic rendering, and composition.
  • Efficient Inference: Quantization reduces memory requirements while maintaining high image quality.

3. Qwen-Image 2.1 Text-to-Image / Image Editing (Qwen-Image 2.1 INT8)

A dedicated INT8 ComfyUI workflow for Qwen-Image 2.1, optimized for efficient image generation and image editing.

  • INT8 Inference: Uses the INT8 version of Qwen-Image 2.1 for reduced VRAM usage and efficient inference.
  • Sampling Preview: Provides an intermediate sampling preview for monitoring the generation process.
  • Face Detailer: Adds dedicated face-detail refinement after generation.
  • Mask Editing: Enables localized image editing by specifying the target region with a mask.
  • Image Editing: Supports Qwen-Image 2.1 image-editing capabilities.
  • Reference Images: Supports reference-based image generation and editing.
  • Native RGBA: Supports native RGBA image generation and transparent backgrounds.
  • Ollama Prompt Optimization: Uses an Ollama-based prompt rewriting workflow optimized around Alibaba's official Qwen-Image prompt-rewriting skills.
  • Official Prompt Skills: The prompt optimization follows the structure and principles of Qwen-Image 2.1's official prompt-rewriting system prompts for Text-to-Image and Image-to-Image editing.

Qwen-Image 2.1:
https://huggingface.co/Comfy-Org/Qwen-Image-2.1

Qwen-Image 2.1 Project:
https://github.com/QwenLM/Qwen-Image-2.1

Supported editing workflows include:

  • Text-to-Image
  • Image-to-Image
  • Mask-based local editing
  • Face Detailer
  • Reference image editing
  • Native RGBA / transparent image generation

4. YuE2 Music Generation (YuE2)

A dedicated ComfyUI workflow for YuE2-3B, an open music-generation model from the Multimodal Art Projection (m-a-p) team.

YuE2 generates complete songs from a style prompt and lyrics, producing vocals and instrumental accompaniment.

  • Text-to-Music: Generate complete songs from style and lyrics.
  • Vocals & Instrumentation: Produces full songs with vocals and accompaniment.
  • Symbolic Planning: Uses symbolic musical planning for melody and harmony.
  • Complete Song Generation: Designed for full-song generation rather than short audio clips.
  • High-Quality Audio: Supports high-quality stereo music output.

YuE2 Model:
https://huggingface.co/Comfy-Org/YuE2

YuE2 Project:
https://github.com/multimodal-art-projection/YuE

Typical input structure:

Style:
J-pop, female vocal, energetic electronic rock, bright synths, live drums

Lyrics:
[Verse]
...

[Chorus]
...
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support