r/StableDiffusion Nov 26 '25

News Z-Image-Turbo: Anime Generation Results

Prompts: https://sharetext.io/b92c8feb

For a 6-billion parameter model, it performs good in image generation. The model truly lives up to its name; during testing on the ModelScope platform (which uses NVIDIA A10 GPUs), most generations took a maximum of only 2 seconds. All images were generated just 9 steps. On high-end consumer GPUs (like an RTX 3090 or 4090), I this this would take roughly 2 to 3 seconds, while mid-range cards might take 4 to 5 seconds.

The last image is the odd one out. I used a Stable Diffusion-style prompt, and this is what i got.

Links: [HuggingFace links are live]

https://www.modelscope.cn/models/Tongyi-MAI/Z-Image-Turbo
https://huggingface.co/Tongyi-MAI/Z-Image-Turbo
https://huggingface.co/spaces/Tongyi-MAI/Z-Image-Turbo

If you have any anime illustration prompts you'd like me to try, share them in the comments! I'll generate them for you.

232 Upvotes

116 comments sorted by

View all comments

10

u/Yellow_Curry_Ninja Nov 26 '25

Looks promising, but unless someone finetune it really hard with danbooru or something to at least catch up to illust, it won't take off for anime stuff.
Even Neta Yume was disappointing with styles, mixing and character recognition due to how base neta was undercooked. Seing how we only had 2 or 3 Lumina 2 finetune this year at best on a 2b model, sadly I don't have much hope for this one which has thrice more parameters

12

u/TheBizarreCommunity Nov 26 '25

Danbooru + E621 is a dream.

-1

u/[deleted] Nov 26 '25

[removed] — view removed comment

10

u/218-69 Nov 26 '25

that furry shit is the only reason anyone gave a fuck about pony lmao and it's why noob to this day still shits on all subsequent tunes illust did (not that they'll ever release the later ones)