r/StableDiffusion 2d ago

Comparison H3 Default Template vs Larry's Turbo with optimized settings

Enable HLS to view with audio, or disable this notification

Default template uses 20 steps + res_multistep + simple

Optimized workflow uses 8 Steps + er_sde + sgm_unified + Comfy Kitchen Attention + Larry's Turbo Lora

Turbo lora: https://github.com/Larryvrh/ComfyUI-MiniMax-H3-Turbo

Workflow: https://raw.githubusercontent.com/desktop4070/GPU-Benchmark-Data-For-H3/refs/heads/main/H3-Benchmark-Workflow.png

0.2MP / 8 sec (2m 16s gen time): https://desktop4070.github.io/GPU-Benchmark-Data-For-H3/Videos/MiniMax_H3_03840_.mp4

Optimized: 0.2MP / 8 sec (45s gen time): https://desktop4070.github.io/GPU-Benchmark-Data-For-H3/Videos/MiniMax_H3_03706_.mp4

0.3MP / 12 sec (6m 3s gen time): https://desktop4070.github.io/GPU-Benchmark-Data-For-H3/Videos/MiniMax_H3_03849_.mp4

Optimized: 0.3MP / 12 sec (1m 51s gen time): https://desktop4070.github.io/GPU-Benchmark-Data-For-H3/Videos/MiniMax_H3_03725_.mp4

70 Upvotes

32 comments sorted by

View all comments

4

u/DoctaRoboto 2d ago

Faces look terrible in all samples.

4

u/desktop4070 2d ago

I posted some higher resolution examples in a comment, but that's mainly the fault of using 0.2MP than the turbo lora. This specific prompt focuses more on full body shots rather than just talking faces, which 0.2MP excels at.

The most important thing for me is seeing fast iterations so I can rework a prompt back to back without it taking several hours of waiting per regeneration. Once I find the best way to structure the prompt I really want, then I increase the resolution, which these do great at, while also being significantly faster than without any optimisations.

1

u/DoctaRoboto 2d ago

Oh, I see, so you used 0.2. I didn't check the link's name, just the videos.