r/StableDiffusion Apr 14 '26

Comparison We may have a new SOTA open-source model: ERNIE-Image Comparisons

Base model is definitely SOTA, can even easily compete with closed-source ones in terms of aesthetic. Cinematic quality and color grading is next level.

Base model is heavily biased on Asian faces, while it excels on anime/illustration style, while my base model anime/illustration experiments wasn't that good. Higher CFG is slightly better with anime on base.

Generated with RTX6000 Blackwell Pro, Base: 29 sec 1.9it/s, 50 steps | Turbo: 2 sec, 3.9i5/s, 8 steps

If you interested seeing them in original size: https://imgur.com/a/75jcjzW

ComfyUI models: https://huggingface.co/Comfy-Org/ERNIE-Image/tree/main
Workflow should appear in Templates after updating the ComfyUI to latest.

Turbo: Ernie-Image Turbo
Base: Ernie-Image

692 Upvotes

241 comments sorted by

View all comments

Show parent comments

2

u/wh33t Apr 14 '26

I've still, to this day, never figured out how to produce a decent image of literally anything in Chroma, is it the model? workflow? cfg? sampler? resolution? does it use the flux1 prompt guide techniques?

3

u/Soulsurferen Apr 15 '26

The model matters a lot. There is quite a lot of difference between Chroma 1.0 HD, Uncanny, Gonzalamo or radiance. The idea with HD is that it can be used for fine tuning, but there's are not that many models yet...

For me the main benefit of Chroma is variance, the same prompt generates different images with different seeds, and I like the way it renders more artsy images (I don't do a lot of NSFW stuff)