LLaDA-Image Turbo: 4-Step Generation and Editing by InclusionAI

ComfyUI Wiki

LLaDA-Image Turbo distills InclusionAI's 6B unified image model to 2-4 steps with Twin-DMD, running text-to-image and native instruction editing in ComfyUI.

L

LLaDA-Image Turbo

Text-to-ImageImage EditingFew-StepDistilled

Twin-DMD distilled variant of InclusionAI's 6B unified image model. Compresses both generation and editing to 2 to 4 sampling steps while keeping the native editing path, and ships as BF16 and INT8 transformer weights for ComfyUI.

DeveloperInclusionAI (Ant Group)
Release Date2026-09-04
ArchitectureUnified diffusion backbone + DiT
Parameters6B
LicenseApache-2.0
Inference2-4 steps, CFG 1.0

Capabilities

  • Text-to-image generation: 50 steps on the Base model, 2 to 4 steps on Turbo
  • Instruction-guided editing: uses the model's native editing path with reference preservation, not a denoise-strength img2img approximation. Edit dimensions must be divisible by 32
  • Bilingual text rendering: legible Chinese and English text in generated images
  • VQ-conditioned generation: SigVQ conditioning module

ComfyUI Usage

The community RebelAI node pack provides four nodes: LLaDA Image Loader, LLaDA Image Text to Image, LLaDA Image Edit, and LLaDA Image Unload.

Place the optimized weights in their ComfyUI folders:

FileDestination
LLaDA-Image-Turbo-transformer-BF16.safetensors or LLaDA-Image-Turbo-INT8.safetensorsComfyUI/models/diffusion_models/
LLaDA-Image-Turbo-text_encoder-Q4_K_M.ggufComfyUI/models/text_encoders/
LLaDa_VAE.safetensorsComfyUI/models/vae/

The INT8 transformer is a native safetensors quantization; GGUF is used only for the quantized LLaDA2-MoE text encoder. Start with 4 steps and CFG 1.0.

Guides and workflows related to this model series.

No articles found.

Comments

Sign in with GitHub to join the discussion.

Loading comments…