Updated: Sep 1, 2026
styleDownload
2 variants available
SafeTensor
380 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
(2)
Sep 1, 2026
Initial Aria ZIB release with two downloadable diffusion-transformer files.
BF16 Quality: start at 28 steps, CFG 4.0, standard Z-Image Base FlowMatch Euler.
Turbo INT8 Metal: start at 8 steps, CFG 1.0, Euler, simple scheduler, model sampling shift 3.0.
The Aria cinematic LoRA is already fused into both files. The INT8 version also has Alibaba-PAI 2603 8-step distillation fused at 0.8. Do not load either LoRA again.
Requires a separate Z-Image-compatible Qwen3 4B text encoder and ae.safetensors VAE. Tested locally on Apple M4 Pro / 24 GB. No on-site generation configuration or BF16-parity claim.
Show more

7690 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
20 1 2 3 4 5 6 7 8 9
License:
Apache 2.0ARIA ZIB
Aria ZIB is a downloadable Z-Image Base derivative built for cinematic photography and stylized film stills: rich but controlled color, organic surface texture, soft halation, volumetric light, and deliberate photographic composition.
This release is presented and validated as an SFW cinematic model. It was not developed or validated as an explicit-anatomy or NSFW-specialist checkpoint.
TWO DOWNLOADS, ONE ARIA LOOK
BF16 Quality — reference-quality rendering and the full Z-Image Base CFG workflow. Start with 28 steps and CFG 4.0.
Turbo INT8 Metal — a faster, lower-memory option for a compatible Metal/INT8 stack. Start with 8 steps, CFG 1.0, Euler, simple scheduler, and model sampling shift 3.0.
No additional Aria LoRA is required. The Aria cinematic adaptation is already baked into both files. The INT8 version also contains the 2603 8-step distillation at strength 0.8; do not load that distillation LoRA again.
REQUIRED SEPARATE COMPONENTS
These downloads contain the Z-Image diffusion transformer only. Install a compatible Qwen3 4B text encoder and the Z-Image ae.safetensors VAE separately, then load Aria through ComfyUI's Load Diffusion Model / UNETLoader in a normal split-component Z-Image Base workflow.
OPTIONAL STYLE TOKEN
aria_cinema
The token can reinforce the intended house style, but it is optional.
HONEST LINEAGE
Aria ZIB was not trained from scratch and is not a full 6B-parameter fine-tune. It starts from Tongyi-MAI Z-Image Base. The owner trained a rank-32, alpha-16 cinematic LoRA for 600 optimizer updates on 299 owner-generated reference images, then fused the final adapter into Z-Image Base at strength 1.0.
BF16 Quality is the fused Aria model in native ComfyUI Z-Image format.
Turbo INT8 Metal additionally fuses Alibaba-PAI Z-Image-Fun 2603 8-step distillation at strength 0.8, then applies INT8 ConvRot quantization while retaining sensitive matrices in BF16. It is not the official Z-Image-Turbo model, and no BF16-parity claim is made.
The training reference set consisted of images generated locally in ComfyUI by the model owner using Krea 2 / Krea 2 Turbo-family community derivative checkpoints, with optional third-party LoRAs in parts of the source archive. Those tools rendered reference images only: no Krea-family checkpoint or LoRA weights were copied or merged into these downloads.
TESTED SETUP
Apple M4 Pro with 24 GB unified memory, ComfyUI 0.31.0, comfy-kitchen 0.2.28, PyTorch 2.13.0, and locally modified Apple-Silicon INT8/FP8 support based on ComfyUI-AppleSilicon-FP8 1.3.1.
Other Macs, NVIDIA/AMD systems, and unmodified plugin builds have not been validated. Loading behavior, speed, memory use, and output can vary by backend. This release does not configure or promise Civitai on-site generation.
KNOWN LIMITATIONS
• Small or lengthy text inside images can be malformed.
• Complex human anatomy and close-contact poses can fail.
• The cinematic adaptation can be visually strong and varies with subject, seed, aspect ratio, and settings.
• The distilled INT8 variant can change composition or reduce fine detail relative to BF16.
• Third-party LoRA compatibility has not been comprehensively tested.
RELEASE TERMS
Aria ZIB v1.0 is a free, non-commercial community release. There is no paid access, per-generation fee, donation goal, or tip request. Credit is required. Commercial use, resale, paid generation-service use, and distribution of merges or derivatives are not granted by this release page. Upstream components remain subject to their own licenses and notices.
CREDITS
Tongyi-MAI — Z-Image Base
Alibaba-PAI — Z-Image-Fun-Lora-Distill
Hugging Face Diffusers
ComfyUI
Krea 2 — referenced only for the lineage of locally generated training outputs; no Krea weights were merged
