Download
1 variant available

937
License:
Qwen Image 2.1 7B INT8 and INT4 (W4A8) ConvRot Quantizations, for use in ComfyUI
Qwen-Image-2.1
Qwen-Image-2.1, A unified text-to-image generation and image editing model in the Qwen family. With just 7B parameters in its visual generation component (32 Single-Stream DiT layers), Qwen-Image-2.1 balances generation quality, inference efficiency, and versatility.
Highlights
Efficient Image Generation: Qwen-Image-2.1 combines strong visual performance with fast inference and a compact design, making high-quality image creation accessible across a wide range of creative workflows.
Flexible Creative Control: With support for diverse inputs, outputs, and localized edits, Qwen-Image-2.1 gives creators the flexibility to explore ideas and refine details within a unified workflow.
Four key improvements define this release:
Compact and Efficient — A lightweight architecture with mixed-granularity attention and prefix KV cache reuse delivers strong image quality at low computational cost.
Native Transparency, Unified Creation and Editing — Generate regular or transparent (RGBA) images from text, edit transparent layers, and extract subjects from photographs—all in one model.
Versatile Editing — Support up to 10 reference images, specify local edits via circles, painted annotations, or separate masks, and preserve identity for people and products.
Realistic Textures and Refined Aesthetics — Improved typography, portrait lighting, and fine details for more visually compelling results.

