The Latest AIs, every day
AIs with the most favorites on Toolify
AIs with the highest website traffic (monthly visits)
AI Tools by Apps
Discover the Discord of AI
AI Tools by browser extensions
GPTs from GPT Store
Discover The Best Model For AI
Top AI lists by month and monthly visits.
Top AI lists by category and monthly visits.
Top AI lists by region and monthly visits.
Top AI lists by source and monthly visits.
Top AI lists by revenue and real traffic.
Full Fine-Tune โข Rich Aesthetics โข Strong Diversity โข Full Negative Prompt Support
BF16 & FP8 & GGUF & AIO โข Natural Language Prompts โข 8GB VRAM
|
|
|
|
|
|
|
|
|
Z-Anime is a full fine-tune of Alibaba's Z-Image Base architecture โ not a LoRA merge , but a fully trained anime-focused model family built from the ground up.
Built on the S3-DiT (Single-Stream Diffusion Transformer, 6B parameters) , Z-Anime inherits the strong foundation of Z-Image Base: rich diversity, strong controllability, full negative prompt support, and a high ceiling for fine-tuning โ now adapted for anime-style generation.
This repository contains the full Z-Anime family :
| Variant | Focus | Best For |
|---|---|---|
| ๐ Z-Anime Base | Highest quality | Final renders, full control |
| โก Z-Anime Distill-8-Step | Speed + quality balance | Everyday generation |
| ๐ Z-Anime Distill-4-Step | Maximum speed | Fast iteration, batches |
| ๐ฆ GGUF Variants | Lower memory usage | Low VRAM / CPU / AMD-friendly workflows |
| ๐ฆ AIO Variants | Single-file convenience | Easy ComfyUI setup |
| ๐ Diffusers Folder |
from_pretrained()
ready
|
Python pipelines, further fine-tuning |
Full fine-tune on Z-Image Base โ BF16 & FP8
BF16 & FP8 โ fast anime generation in 8 steps , CFG 1.0
BF16 & FP8 โ ultra-fast anime generation in 4 steps , CFG 1.0
Available for low VRAM , CPU inference , and AMD-friendly workflows.
All-in-one checkpoints with
image model + VAE + Text Encoder integrated
in a single file.
Available for
Base
,
Distill-4-Step
and
Distill-8-Step
โ each in
BF16 & FP8
.
The required
VAE
(
ae.safetensors
) and
Text Encoder
(
qwen_3_4b.safetensors
) are also included in this repository for users running the standard (non-AIO) variants.
The full
Diffusers-format folder
(
diffusers/
) is included โ drop-in compatible with
ZImagePipeline.from_pretrained()
for Python users who want to run inference outside ComfyUI or use Z-Anime as a starting point for further fine-tuning.
More updates coming โ follow to stay notified! ๐
Maximum precision. BFloat16 format with minimal quality compromise. Best for final renders, careful work, and LoRA training.
Recommended for most users. Smaller files, faster downloads, and excellent quality with only minor tradeoffs compared to BF16.
Optimized for lightweight inference setups, especially useful for low VRAM, CPU inference, or alternative backends.
All-in-one checkpoints with image model + Text Encoder + VAE integrated into a single file for the easiest setup. Available for Base, Distill-4-Step and Distill-8-Step.
The foundation of the Z-Anime family.
A full fine-tune with the highest quality ceiling , the widest creative range , and full negative prompt support .
steps: 28-50
cfg: 3.0-5.0 # up to 9.0 possible
sampler: euler_ancestral
scheduler: beta
negative_prompt: strongly recommended
Negative prompts have full effect on Z-Anime Base and are highly recommended.
The sweet spot of the family.
Distilled from Z-Anime Base, this version delivers strong anime results in just 8 steps while keeping most of the quality.
steps: 8
cfg: 1.0 # max ~1.5
sampler: euler_ancestral
scheduler: beta
negative_prompt: limited effect
Negative prompts have only limited effect at this distillation level. If your workflow includes ConditioningZeroOut , prefer that instead of a large negative prompt.
The fastest Z-Anime variant.
Built for maximum throughput โ ideal for rapid prototyping, quick batch generation, and speed-focused workflows.
steps: 4
cfg: 1.0 # max ~1.5
sampler: euler_ancestral
scheduler: beta
negative_prompt: limited effect
| Use Case | Resolution |
|---|---|
| Portrait / character art | 832 ร 1216 |
| Landscape / scenes / backgrounds | 1216 ร 832 |
| Square / general purpose | 1024 ร 1024 |
| Tall / full body / wallpaper | 768 ร 1344 |
| Cinematic / wide scenes | 1920 ร 1088 |
| Detailed portraits | 1024 ร 1536 |
Supported range:
approximately
512 ร 512 to 2048 ร 2048
, any aspect ratio.
All main variants are designed to run on
8GB VRAM
.
Natural language works best โ not tag lists.
A young anime girl with long silver hair and golden eyes, wearing a traditional shrine maiden outfit with white haori and red hakama. She stands in a sunlit bamboo forest, cherry blossoms falling softly around her. Warm afternoon light filtering through the trees, detailed fabric shading, expressive face, calm serene expression, high quality anime illustration with fine line work.
anime girl, silver hair, shrine maiden, bamboo, cherry blossom, warm light
Detailed anime portrait of [character], soft rim lighting, expressive eyes with detailed reflections, fine hair strands, clean linework, professional anime illustration quality.
Dynamic anime [scene], dramatic angle, motion energy, speed lines, particle effects, cinematic composition, detailed shading, high quality anime art.
Anime [location] at [time of day], [lighting], [atmosphere], beautiful background art, wallpaper quality, highly detailed environment.
Choose between:
ComfyUI/models/diffusion_models/
โโโ z-anime-base-bf16.safetensors
โโโ z-anime-base-fp8.safetensors
โโโ z-anime-distill-8step-bf16.safetensors
โโโ z-anime-distill-8step-fp8.safetensors
โโโ z-anime-distill-4step-bf16.safetensors
โโโ z-anime-distill-4step-fp8.safetensors
ComfyUI/models/unet/
โโโ z-anime-base-q8_0.gguf
โโโ z-anime-base-q4_k_s.gguf
Two text encoders are included โ pick one :
ComfyUI/models/clip/
โโโ qwen_3_4b-bf16.safetensors # default (Z-Image standard, BF16)
or
โโโ qwen_3_4b-fp8.safetensors # default (Z-Image standard, FP8)
or
โโโ qwen_3_4b-engineer-v4-bf16.safetensors # alternative (Engineer V4, BF16)
or
โโโ qwen_3_4b-engineer-v4-fp8.safetensors # alternative (Engineer V4, FP8)
qwen_3_4b-*
)
โ the standard Z-Image text encoder, repackaged as a single
.safetensors
file (BF16 + FP8). This is what the model was trained against.
qwen_3_4b-engineer-v4-*
)
โ an alternative full fine-tune of the Z-Image text encoder by
BennyDaBall
, drop-in compatible. Often produces more varied outputs from the same seed. See
Credits
below for the original repo.
ComfyUI/models/vae/
โโโ ae.safetensors
For the AIO versions, you only need the single checkpoint file โ no extra VAE or Text Encoder required:
ComfyUI/models/checkpoints/
โโโ z-anime-base-aio-bf16.safetensors
โโโ z-anime-base-aio-fp8.safetensors
โโโ z-anime-distill-8step-aio-bf16.safetensors
โโโ z-anime-distill-8step-aio-fp8.safetensors
โโโ z-anime-distill-4step-aio-bf16.safetensors
โโโ z-anime-distill-4step-aio-fp8.safetensors
Use:
models/unet/
folder
Use a standard Checkpoint Loader โ no extra CLIP or VAE loading required.
For Python users, the full Diffusers-format folder is included under
diffusers/
and can be loaded directly with the
subfolder
argument:
import torch
from diffusers import ZImagePipeline
pipe = ZImagePipeline.from_pretrained(
"SeeSee21/Z-Anime",
subfolder="diffusers",
torch_dtype=torch.bfloat16,
).to("cuda")
image = pipe(
prompt="A young anime girl with long silver hair and golden eyes, "
"shrine maiden outfit, sunlit bamboo forest, cherry blossoms, "
"professional anime illustration, fine line work.",
num_inference_steps=40,
guidance_scale=4.0,
).images[0]
image.save("z-anime-output.png")
This format is also a clean starting point for further fine-tuning (LoRA or full fine-tune) with frameworks like OneTrainer , diffusers , or kohya-ss .
A ready-to-use ComfyUI workflow that supports
all variants
(Base / Distill-8 / Distill-4, BF16 / FP8 / GGUF / AIO) is included in
workflows/Z-Anime-Workflow-v1.json
.
It includes:
Z-Anime/
โโโ README.md
โโโ config.json
โ
โโโ diffusion_models/
โ โโโ z-anime-base-bf16.safetensors
โ โโโ z-anime-base-fp8.safetensors
โ โโโ z-anime-distill-8step-bf16.safetensors
โ โโโ z-anime-distill-8step-fp8.safetensors
โ โโโ z-anime-distill-4step-bf16.safetensors
โ โโโ z-anime-distill-4step-fp8.safetensors
โ
โโโ gguf/
โ โโโ z-anime-base-q8_0.gguf
โ โโโ z-anime-base-q4_k_s.gguf
โ
โโโ aio/
โ โโโ z-anime-base-aio-bf16.safetensors
โ โโโ z-anime-base-aio-fp8.safetensors
โ โโโ z-anime-distill-8step-aio-bf16.safetensors
โ โโโ z-anime-distill-8step-aio-fp8.safetensors
โ โโโ z-anime-distill-4step-aio-bf16.safetensors
โ โโโ z-anime-distill-4step-aio-fp8.safetensors
โ
โโโ text_encoder/
โ โโโ qwen_3_4b-bf16.safetensors # default
โ โโโ qwen_3_4b-fp8.safetensors # default
โ โโโ qwen_3_4b-engineer-v4-bf16.safetensors # alternative (BennyDaBall)
โ โโโ qwen_3_4b-engineer-v4-fp8.safetensors # alternative (BennyDaBall)
โ
โโโ vae/
โ โโโ ae.safetensors
โ
โโโ diffusers/
โ โโโ model_index.json
โ โโโ scheduler/
โ โโโ tokenizer/
โ โโโ text_encoder/
โ โโโ transformer/ (sharded safetensors + index)
โ โโโ vae/
โ
โโโ images/
โ โโโ cover.png
โ โโโ workflow-cover.png
โ โโโ workflow-overview.png
โ โโโ 1.png
โ โโโ 2.png
โ โโโ 3.png
โ โโโ 4.png
โ โโโ 5.png
โ โโโ 6.png
โ โโโ 7.png
โ โโโ 8.png
โ โโโ 9.png
โโโ workflows/
โโโ Z-Anime-Workflow-v1.json
ae.safetensors
) and
Text Encoder
(
qwen_3_4b.safetensors
) included
Tongyi-MAI/Z-Image
BennyDaBall/Qwen3-4b-Z-Image-Engineer-V4
โ full fine-tune with SMART training, included as alternative text encoder
Z-Anime is an experimental anime-focused model family built to explore what a full fine-tune on Z-Image Base can achieve in this space.
It is already strong for anime aesthetics, character work, and fast iteration, and future versions will continue to improve diversity, character handling, prompting flexibility, and overall quality.
Z-Anime โ anime at its finest, powered by Z-Image Base. ๐
Z-Anime huggingface.co is an AI model on huggingface.co that provides Z-Anime's model effect (), which can be used instantly with this SeeSee21 Z-Anime model. huggingface.co supports a free trial of the Z-Anime model, and also provides paid use of the Z-Anime. Support call Z-Anime model through api, including Node.js, Python, http.
Z-Anime huggingface.co is an online trial and call api platform, which integrates Z-Anime's modeling effects, including api services, and provides a free online trial of Z-Anime, you can try Z-Anime online for free by clicking the link below.
Z-Anime is an open source model from GitHub that offers a free installation service, and any user can find Z-Anime on GitHub to install. At the same time, huggingface.co provides the effect of Z-Anime install, users can directly use Z-Anime installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

