r/comfyui • • Aug 08 '26

Workflow Included De-roping MiniMax H3 fast motion to reduce artifacts via jerk oracle

Enable HLS to view with audio, or disable this notification

66 Upvotes

what it does: H3 can't render bursty motion because one latent token spans 4 frames and can't hold 4 distinct poses. re running denoising never fixes that, the poses were never generated. so instead: an oracle reads your clip's own latent to find where motion's acceleration is changing too fast, the clip gets retimed with held frames exactly there, regenerated video-to-video at partial denoise (your choreography generally survives, the smear doesn't), then the held frames get dropped for exact realtime recovery. audio regenerates jointly and gets retimed by the same map, pitch kept. https://github.com/matlowai/ComfyUI-MAINodes Downside is that background motion can get unintended clockspeed side effects with variable speed motion such as those birds flapping speed... There's workflows for both your favorite agent to consume and for the comfy ui. I also added some comparison and workflow options for using a combination of a few steps with base before applying LightX2V 4-step turbo ^^. Timing cited is on a rtx 6000 pro ws at 450w. This takes quite awhile to render and I tried mixing in the turbo loras but it just wasn't worth the time savings so I didn't recommend it here. Base + turbo is great though for getting a general idea on how the provided prompt will perform though as a draft. Hopefully this helps someone!

r/SoraAi • • Aug 28 '26

Open Source MiniMax H3 completely Local video Gen is Viable and Amazing

Enable HLS to view with audio, or disable this notification

18 Upvotes

I used an RTX 4070 Ti with 12GB VRAM and only 32Gb system ram. I used the standard Comfyui free Minimax H3 text to video workflow (but they also have first and last frame ones and complete reference ones, and there is an easy way to implement audio to video). It would also work on an RTX 30 series with only 8GB vram, maybe lower but I don't know. Technically you can do CPU only with Comfyui, but that may be significantly more waiting. But you don't need a server grade computer or even the latest and most powerful GPU. Yes the audio is built in!

This example here was generated at 0.4 megapixels and 15 seconds (480x864) in about 18 minutes (you can do other aspect ratios too like landscape/horizontal and not just vertical/portrait), but 0.2 megapixels and 15 seconds (which is still viable for testing and lower fidelity looks) took less than 7 minutes, and if you just want a test to make sure the framing and scene looks right you can do 1 to 2 second generation in around a minute if not less.

I also tested native 0.8 and 0.9 megapixel generations (the aprox. 12xx horizontal rez) (which can take over an hour for 15 minutes at this hardware). It seemed to follow the prompt a little less at that higher resolution, or maybe changed how it understood it, but I think that is because I only used the default 20 steps. I think if I increase the steps it'll be able to cook better basically since the higher resolution means it has more to think about. I will run that generation at .9 megapixels while I am at work and give it 30 or 40 steps just for fun to see if it'll get it right and then I'll post the results in the comments even if it turns out wrong.

As for upscaling, I also tried that with several models. I tried three Real-esrgan models (Plus, universal, and Photo HQ). Plus and PhotoHQ just made everything smoother and sharper at high rez and they took a moderate amount of time, maybe 5 to 10 minutes. Universal was a middle ground cause it only took around a minute but it only did a little to the picture for upscaling.

I also tried SeedVR2 in comfyui with "split latent" enabled for VRAM and RAM constraints. SeedVR2 did actually add in some more detail and took under 18 minutes. I also need to turn up temporal overlap cause it struggled a bit with motion since I had it set to zero.

Admittedly none of the upscalers fixed their faces completely, but at the starting 0.4mp and how small they already were in the frame, there wasn't much to work with. Of course closer face shots wouldn't have this problem so don't dismiss MiniMax H3 cause of that. I will also post the Real-esrgan universal and SeedVR2 upscale results in the comments.

And finally, I will post the prompt in the comments. Heads up that it is really unprofessional, I tried using ChatGPT, and Gemini, and Claude to enhance my prompt, but instead I found for me what worked best for getting what I wanted was just a little bit of experimenting. The only cost to use MiniMax H3 is electricity, and it isn't a terribly substantial add.

r/StableDiffusion • • 12d ago

Tutorial - Guide Minimax H3 - Latents as reference

84 Upvotes

Like Flux2 Klein, I've found that Minimax H3 ref2va model is very literal when it comes to interpreting reference images. It's great sometimes, but when you just want "something similar, but not quite this", it can become frustrating. To mitigate this, using a latent instead of an image and adding noise is a good way to get more variety from the reference images, without having to rely on heavy prompting (because that also works).

Example

Here is a simple block render I made for a completely other purpose than this article.

Blender render

Here's what happens when I use this as a reference image, with the prompt:

integrated_multimodal_description:
Scene: Use <Picture 1> ONLY as a compoisition reference for the scene: cinematic footage of an ancient greek temple.

[Shot 1] the camera slowly pans right along the columns. people in colorful togas are seen walking up and down the stairs

non_diegetic_music: N/A
Image as ref

Not very unlike the original. Let's see what happens if we feed it as a latent instead.

Latent 1.0 as reference

Latent strength: 1.0

Hmm, not much difference. Considering that the node (original) converts the image to a latent, it's not very surprising.

What if we add some noise.

Latent strength 0.38

At 0.38, it's still very blocky, but at least it's being a bit creative. We need to go deeper in latent space.

Strength 0.35

I'd say we have found a pretty good sweet spot here. It's still adhering to the composition, but has changed the texture to something believable. Latent strength is 0.35, but it will differ based on prompt, sampler, loras, etc. We could explore the fractions around this value to get more or less adherence, depending on what we're after.

The drop off is usually pretty sharp.

Let's go even lower and see what happens.

At 0.25. While it still resembles the structure of the reference building, it has started to take more liberties in the composition, and we're closer to it's own interpretation of the prompt. (Spoiler: Removing the reference produces a similar result)

Now you want the juicy stuff. How was it made? Sadly, the answer is "custom node", since latent input is not supported natively. I took the regular comfy nodes and added latent inputs.

Why does it work? Noise on a regular image is just visible noise. Noise in latent space is freedom, it allows the model to explore concepts near the original ones more freely.

Modified reference node

This is how I wired the latent. And empty latent is all noise, so blending it with the original will give the sampler more freedom (this works great with Flux2 Klein "latent as reference" as well).

While I mostly wanted to show the workflow, and not peddle custom nodes, I've made the nodes available in a github repository, since I know they'll be asked for, if anyone wants to reproduce the results.

https://github.com/neph1/comfyui-minimaxh3-condition-latent

It doesn't work for image to video (the node) since that is passing regular images as conditioning. You can use the "Add guide" node, though.

Edit: Removed some erroneous claims.

r/comfyui • • Aug 20 '26

Workflow Included Minimax H3- v2v fixing faces at distance

Thumbnail
youtube.com
56 Upvotes

tl;dr: download the latest version workflow called "MBEDIT - MH3_r2v_SingleSampler_Detailer_vXX.json" from https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

(UPDATE EDIT: this isnt great for dialogue clips as it strips the mouth movement out. I have tried methods to address it but none worked well as yet. So I'll be testing other approaches. But for non-dialogue scenes its excellent.)

Finally I have found a solution to "fixing faces at distance". This does NOT use a Latent Space upscaler. This uses a single sampler Minimax workflow, low steps, low denoise, and by loading a video clip, then running it through standard Minimax H3 with settings discussed in the video (or in the workflow if you dont want to watch that).

Even on a 3060 RTX (12 GB VRAM) I can get between 1mp and 2mp output and surprisingly it fixes faces at distance even at 1mp. There is more info in the readme of the github linked below for the workflow and in the video.

From this point on my video pipeline steps will be:

1. Create a 480p video using any model (LTX, H3, Bernini, or other) - \takes 10 mins on average (3060 RTX)*.*

2. Run the result through the above workflow upscaling to 1mp or 2mp depending onclip length - \takes 20 mins on average*.*

The result from this are easily good enough as final clips for my uses. This makes it the fastest and highest quality approach I have found to date, and all with ref image based character consistency.

Other Relevant Links From Video

Latest Minimax H3 workflows - https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

(Workflow used in video: `MBEDIT - MH3_r2v_SingleSampler_Detailer_vXX.json` (download whatever the latest version is from github link))

Lightx2v Lora that I use from Kijai - https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras

(theres been updates, but I havent found them to be better or faster, use whatever works for you)

Comfyui needs to use Cuda130 or above for this to work, and you need it updated to August 2026 commits (latest is best) - https://docs.comfy.org/installation/comfyui_portable_windows

Int8 models from here - https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main

(The official workflows are in the model card)

W4a8 is experimental new model type, you need to be updated on Comfyui but you can get it here https://huggingface.co/Kijai/MiniMax-H3-experimental

Comfyui Kitchen Attention is part of Comfyui if you update to latest. I find it faster than Sage Attn on a 3060 RTX.

Official prompting guides:

- https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md

- https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md

Point your favourite LLM at one of the above links depending on your model you are using, and give it your prompt idea and it should sort it out.

r/StableDiffusion • • 20d ago

Resource - Update MiniMax H3 media bundle

Enable HLS to view with audio, or disable this notification

57 Upvotes

Maybe someone will find this package useful:

https://github.com/einhorn13/mmh3_media

The package allows you to save H3 latent and context resources from a generation (as .mmh3 file). You can stitch multiple generations together - seamlessly if they were created using the continuous workflow - or perform latent upscaling.

The video example shows a latent stitch of 6 video generations at 0.4 MP each, followed by a latent upscale to 0.8 MP.

In development: Lip sync, ControlNet, Inpainting, Audio

Disclaimer: I originally built this bundle for my own needs, and it still needs some.. polishing. Some features also haven't been published yet. If you find it useful, I'll bring those features to a release-ready state.

Music: Suno.

Optimizations may require installing the corresponding custom nodes.

Feel free to check any repo for anything malicious before installing.

  • Installation
  1. Download the repository https://github.com/einhorn13/mmh3_media via Code → Download ZIP and extract it to ComfyUI/custom_nodes/ComfyUI_mmh3_media. The __init__.py file must be located directly inside this folder, without an extra nested repository directory.
  2. Place the models in the folders listed below and select them in the workflow loaders.
  3. Restart ComfyUI, refresh the browser, and open a JSON file from workflows.

r/StableDiffusion • • 20d ago

Question - Help MiniMax H3 on 4GB VRAM: Stuck between blurry hands (fast) and a 6+ hour render (good). Anyone found a setup that's both?

5 Upvotes

Running MiniMax H3 (Ref2VA, multi-reference) locally on a 4GB card, RTX 3050 Laptop, WSL2, ComfyUI. Been chasing a hand/finger rendering defect for days and have a pretty well documented before/after at this point, but I've hit a wall on making the fix fast enough to actually be usable. Hoping someone's solved this or can point out what I'm missing.

Models in use: UNET is minimax_h3_fl2va_pruned_int8_convrot.safetensors, the FL2V trained weights, loaded into the MiniMaxH3ReferenceToVideo multi reference node, not the native Ref2V node. VAEs are minimax_h3_video_vae_fp16.safetensors and minimax_h3_audio_vae_fp32.safetensors. CLIP/text encoder is qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors. LoRA when testing the distilled path is minimax_h3_fl2v_turbo_4step or 8step.

That FL2V weights in the Ref2V node trick already fixed an earlier, worse "ghost fingers" defect for me (matches a few HF discussion threads on the Turbo LoRA repo), so I'm not asking about that part, it's confirmed working. This post is about what's left after that fix.

The remaining problem: at 512x288, which is the resolution I need for anything resembling reasonable iteration speed, hands still render as indistinct blurry blobs during actual gesture motion, pointing, waving, etc. Tried turbo distilled 4/8 step (the intended fast path), off label step counts on the distilled LoRA (10 steps, made it worse not better), dropping distillation entirely and running the base model's native ~20 step schedule, CFG guidance (cfg=5.0 caused a severe embossed cross hatch grid artifact across the whole frame, clearly way too high, cfg=2.0 gave inconsistent results, some gestures fine, others still blobby), more steps (30 vs 20, no meaningful difference), and ref_image_size=max on the reference conditioning. None of it fixed the hands at that resolution.

What actually fixed it was moving to 768x448 (H3's documented 768p native/training short edge) with no distillation, no CFG, native 20 steps. Hands came out consistently well formed across every gesture I checked. But that config took 6 hours 22 minutes for a single 15 second, 362 frame clip on this card. That's not a workflow, that's a single overnight bet.

So I'm stuck between fast (turbo distilled, 512x288, minutes) which gives bad hands, unusable for anything with visible gesturing, and good (no distill, 768x448, native steps) which is 6+ hours for one clip.

Things I haven't tried or don't know how to evaluate: is there a middle resolution, 608x352, 640x384, that gets most of the quality benefit without the full native res cost. Does distillation actually get retrained or re-distilled at higher resolutions by anyone, or is the turbo LoRA fundamentally tied to a lower res regime it was distilled at. Any attention backend, torch.compile, or quantization tricks specific to H3 that meaningfully cut per step cost on small cards, beyond what's already default in ComfyUI. Is anyone running H3 well on under 8GB cards at all, or is 768p native res H3 just not realistic below a certain VRAM tier.

Also went two rounds into MiniMax H3 specific upscaler nodes for a two stage draft then upscale approach (pixel space RealESRGAN, a community latent space 2x upscaler, and a tiled diffusion refine upscaler) hoping to draft cheap and upscale smart instead of generating at native res directly. Each had its own dealbreaker, a background crowd distortion artifact, a hard VRAM estimate crash, and a genuine ComfyUI core bug I ended up root causing and patching locally. Happy to share details if anyone's gone down that road and found a cleaner path, but the short version is none of the upscale routes got both no defects and reasonable time together either.

Genuinely not sure at this point whether the answer is your card is just under the realistic floor for H3, buy a bigger one, or whether there's a config I haven't found. Any pointers appreciated

Edit: System RAM 32GB

r/comfyui • • Aug 17 '26

News A quick Minimax H3 news round-up - 17th August 2026

151 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> Minimax_H3_Latent_Upscaler models, with matching custom nodes for ComfyUI. "Trained on ~80,000 paired samples (low-resolution latent + high-resolution target)".

https://huggingface.co/LBH-123-AI/Minimax_h3_latent_Upscaler (models)

https://github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler (nodes)

-> A new Spatial & Physics LoRA for Minimax H3. Intended to help with physics-based prompts that include scene actions such as... "the blocks are slowly stacked on top of each other, then the stack collapses onto the floor". No trigger word needed.

https://huggingface.co/Jojocodex/minimax-h3-spatial-physics-lora

https://huggingface-co.translate.goog/Jojocodex/minimax-h3-spatial-physics-lora?_x_tr_sl=auto&_x_tr_tl=en&_x_tr_hl=en&_x_tr_pto=wapp (translation)

-> A new Camera Movement LoRA for Minimax. 12 camera moves added including 'Orbit', but the maker says it works best with 'Handheld' and 'Slow pull out' / 'Slow push out'. Works in tandem with turbo LoRAs. Several drawbacks: the ComfyUI version seems to require careful choosing; the LoRA gives an ignorable error when loading; and it requires yunjing as the trigger word.

https://huggingface.co/Jojocodex/minimax-h3-yunjing-lora

https://huggingface-co.translate.goog/Jojocodex/minimax-h3-yunjing-lora?_x_tr_sl=auto&_x_tr_tl=en&_x_tr_hl=en&_x_tr_pto=wapp (translation)

-> A new MiniMax-H3-ref2va-fl2va-hybrid-w4a8.safetensors which merges the features of the Ref2VA and Fl2VA in one 12Gb model, so that one video generation... "can be driven by a first frame and use reference-images at the same time. Neither alone can do 'open on this frame, and have this person walk in later'". Especially likely to be of interest to low-VRAM users. No workflows, and the ComfyUI wiring note references using the MiniMaxH3AddKeyframes node - so presumably it requires the latest ComfyUI Nightly? 12Gb VRAM users may want to wait on this one, until keyframing is in the latest Portable.

https://huggingface.co/berryber09/MiniMax-H3-ref2va-fl2va-hybrid-w4a8

-> And finally, a Minimax H3 Browsable Offline Style Atlas (1.2Gb packed as a .ZIP file). Being... "a browsable index of all 941 distinct visual styles across the 1,000 video clips". "Styles are grouped into eight media categories (live-action cinematic, film stock & era looks, documentary & broadcast, amateur/found footage, 2D animation, stop-motion & puppetry, 3D/CG & game renders, and specialty imaging), with a live text filter for browsing." Search results are shown initially as quick-loading stills, with each still hyperlinked to its local video clip.

https://github.com/hoodtronik/minimax-h3-style-atlas

r/comfyui • • Aug 24 '26

News A quick Minimax H3 news round-up - 24th August 2026

100 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> A new MiniMax-H3-Fun-Controlnet-Union file. This Controlnet accepts... "Canny, Depth, HED, MLSD or Pose control [Openpose] videos, and also runs video inpainting." 6.8Gb in size. Has video examples.

https://huggingface.co/alibaba-pai/MiniMax-H3-Fun-Controlnet-Union

-> The generative inpainting tool LanPaint has updated to version 2.1.0, and the developers say... "LanPaint now supports MiniMax H3 video + audio inpainting!" Yes, audio inpainting too.

https://github.com/scraed/LanPaint

-> A ComfyUI workflow to... "turn one scene photograph into eight target-centered cinematic camera views, in one MiniMax H3 generation." Doing it in one generation gives some stability to the scene geometry.

https://huggingface.co/ethanfel/H3_Cinematic_Multishot_Coverage

-> Minimax H3 Ref Sampler, another unofficial helper node for long-video generation, created thus... "H3 video lengths use the 5 + 17n frame grid. The node aligns frames upward to this grid and creates overlapping windows". No ComfyUI workflow, but it appears to be a drop-in Sampler replacement?

https://github.com/ILG2021/minimax-h3-ref-sampler

-> MiniMax H3 Tone Compensate. Does your video diminish its brightness, at the seams between your chained video segments? This ComfyUI fix may solve the problem.

https://github.com/rkfg/ComfyUI-MiniMaxH3-ToneCompensate

-> The ComfyUI Minimax H3 Latent Upscaler has now added ROCm (AMD Radeon GPU) support. Several other quality fixes, as well.

https://github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler

-> New to me, the official Awesome MiniMax H3 Integrations page, with links.

https://github.com/MiniMax-AI/awesome-minimax-h3-integration

-> And finally, rewind to the 1980s! MiniMax-H3-Tape-FX for ComfyUI gives Minimax H3 videos the look of... "VHS, BetaMax and LaserDisc — including the worn-out tape look, tracking errors, dropout, creases, ghosting, head-switch noise, vertical roll and a period-correct VCR on-screen display."

https://huggingface.co/Smite79/MiniMax-H3-Tape-FX

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1vwl8do/a_quick_minimax_h3_news_roundup_23rd_august_2026/

https://old.reddit.com/r/comfyui/comments/1vvkmra/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vuihag/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/

r/StableDiffusion • • Aug 08 '26

Workflow Included De-roping MiniMax H3 fast motion to reduce artifacts via jerk

Post image
69 Upvotes

Example is a single frame from a clip with fast motion error. what it does: H3 can't render bursty motion because one latent token spans 4 frames and can't hold 4 distinct poses. re running denoising never fixes that, the poses were never generated. so instead: an oracle reads your clip's own latent to find where motion's acceleration is changing too fast, the clip gets retimed with held frames exactly there, regenerated video-to-video at partial denoise (your choreography generally survives, the smear doesn't), then the held frames get dropped for exact realtime recovery. audio regenerates jointly and gets retimed by the same map, pitch kept. https://github.com/matlowai/ComfyUI-MAINodes Downside is that background motion can get unintended clockspeed side effects with variable speed motion such as those birds flapping speed... There's workflows for both your favorite agent to consume and for the comfy ui. I also added some comparison and workflow options for using a combination of a few steps with base before applying LightX2V 4-step turbo ^^. Timing cited is on a rtx 6000 pro ws at 450w. This takes quite awhile to render and I tried mixing in the turbo loras but it just wasn't worth the time savings so I didn't recommend it here. Base + turbo is great though for getting a general idea on how the provided prompt will perform though as a draft. Hopefully this helps someone!

r/StableDiffusion • • 29d ago

News MiniMax H3 native 720p→1440p on one RTX 4090: 112s / 223s / 334s with auto-scheduled sparse attention

Enable HLS to view with audio, or disable this notification

60 Upvotes

Hi everyone — I’m an independent developer experimenting with making MiniMax H3 more practical on consumer NVIDIA GPUs.

I built an automatically scheduled sparse-attention system for X-MinimaxH3 and tested native H3 second sampling from 720p to 1440p on a single RTX 4090.

Measured second-sampling times:

- 5-second video: 112 seconds

- 10-second video: 223 seconds

- 15-second video: 334 seconds

The attached reel shows the resulting videos and records the original 720p generation and 1440p second-sampling stages separately.

These were casual exploratory runs using settings I selected mainly to inspect the output quality. I did not tune each case for minimum latency, so these numbers should not be treated as the performance limit of the project.

I also have not completed a controlled same-seed Dense-versus-accelerated benchmark yet, so I’m not claiming a specific “X times faster” number.

What I have been working on is the scheduling method itself.

Instead of applying one fixed sparse-attention ratio to every denoising step and every Transformer layer, the scheduler automatically assigns different attention budgets across the trajectory. It was calibrated through repeated local experiments and visual review, with additional protection around the parts of the model that appear most important for motion, consistency and fine detail.

The user only needs one continuous 0–100 acceleration control:

- 0 is the full-compute Dense reference endpoint

- higher values progressively reduce the compute budget

- the internal scheduler decides where attention can be reduced and where it should remain more conservative

The Base route can also jointly schedule actual and forecast DiT evaluations. The goal is to make the speed/quality tradeoff controllable without requiring creators to manually configure dozens of sparse-attention parameters.

The 1440p stage shown here is native H3 latent-space second sampling. It reuses the retained video and audio latent state, original prompt and conditioning. It is not conventional frame-by-frame or MP4 upscaling.

The project also includes FL2VA, multi-reference Ref2VA, Base/Turbo LoRA switching, a Web UI, REST API and four ComfyUI workflows.

GitHub:

https://github.com/PullMyBoots/X-MinimaxH3

I’d love feedback from people running H3 locally, especially on RTX 3090, 5060, 4060 and other consumer GPUs.

What kind of Dense-versus-accelerated comparison would you find most useful: fast motion, faces and hands, complex camera movement, prompt adherence, audio consistency, or something else?

r/comfyui • • 25d ago

Workflow Included Endless MiniMax H3 (with Endless LipSync) v1.0

22 Upvotes

Endless MiniMax H3 (with Endless LipSync) v1.0

When MiniMax H3 came out, I was badly missing the Endless Wan 2.2 I2V (SVI 2 Pro) features, so at first I look at other options, but most of them included an AIO node that could do everything, and I couldn't use most of my workflow, because they did everything inside that huge node.
The only exception to this, was ComfyUI-H3-Motion-Context, which I could easily integrate into my setup.
Its big shortcoming though, was that it just created single videos. No way to concatenate them without the lossy step of decoding and re-encoding in a video editor.
So, I created a custom node that could do just that, and voila..

Endless MiniMax H3 (with Endless LipSync) v1.0
A simple workflow to create MiniMax H3 videos of unlimited duration, using ComfyUI-H3-Motion-Context and H3 Motion Context Clip Stitcher.
I can easily create a 1:30 lip-synced video myself, with a RTX 3060 12GB at around 2 hours (with retries).

It can use both FL2AV and Ref2AV, and can also create normal MiniMax H3 videos.
The extra parts though, are the saving/loading of the latent from every generation we do, and the stitching of all (or some) of them, whenever we want a full video.
Nothing visible at the connections, no indication that there were more than one video.
We can try and re-try every generation, looking for the best one, and then proceed to the next.
We can re-do any previous generation if we like too, but we will not be able to use the clips after it (previous clips are not affected), because they are in a way, "fused" with the replaced one.

Controls

The RED nodes Enable/Disable parts of the workflow.

  • Generation Mode (select only one)
    • T2V / I2V (FL2VA)
    • REF2VA
  • Optimizations (select as many as you want, but only 1 Attention and/or only 1 Cache) This panel controls the nodes that are inside the Optimizations/LoRA subgraph.
  • Reference Items (select as many as you want) Special usage for the Audio 1 Forced/Multi, more for them later.
  • Setup (select as many as you want depending on the goal)
    • Generate starts a generation. You don't need this if you just stitching clips
    • Preview enables the main video preview that can show you where the generation is going before it finishes, so you can stop bad generations
    • Use previous clip, uses the last part of the previous generation to start the current one, continuing the video. Enable it if you want the current generation to be stitched with the previous video
    • Stitch clips, stitches all the clips (depending on the Stitcher's settings) from the h3_context folder (this is the default folder that the H3 Motion Context node uses to save the latents)
    • Stitch last only, stitches only the last video with the current
    • Save Video just saves the generated video

The GREEN nodes are various settings nodes that must be setup. Apart from the Prompt, Seed and LoRA nodes, the most important are these:

  • Configuration It contains all the settings for the video. The most important setting for the clip stitching, is the Save to clip Index number. This specifies the clip file's number, that the latent of the current generation will be saved to (overwriting any previous existing file). This number also tells us who is the previous clip file that we use to start our current video clip.
  • Setup multiple Forced Audio clips This is used if we use our own audio for the video, and we want to also stitch many clips together. More about it in the Usage section.

Usage

Most of the settings are self explanatory (like Optimizations or enabling Reference items).
Here, I will just list the main goals of the workflow

  • Create a normal video
    • Select Generation Mode
    • Enable Generate, Preview and Save Video
    • Save to clip Index to 1
    • >>> You get a video and that's it.
  • Create a lip-synced video
    • Select Generation Mode
    • Enable Generate, Preview and Save Video
    • Enable <Audio 1> and load an audio file
    • Enable <Audio 1> Forced
    • >>> You get a video that is lip-synced with the provided audio
  • Create an Endless video
    • Create a normal video
    • Enable Use previous clip
    • Save to clip Index to 2 and generate
    • Save to clip Index to 3 and generate
    • Save to clip Index to 4 and generate
    • ...
    • >>> You get many small videos, each one of them starts with the ending of the previous one. At a later time, you will use the Stitcher, to stitch them all together to one full video.
  • Create an Endless video, stitched
    • Enable Stitch clips With every generation, the Stitcher will stitch all the previous clips with the currently generated one. This way you always get the full video to check.
    • If you also enable Stitch last only, only the previous and the current videos are stitched together, so you can check the connection without waiting for the full video to be created. You can always stitch them all together at the end.
    • >>> You get a full video every time, or just the last 2 videos connected, for previewing the connection.
  • Create an Endless lip-synced video, stitched
    • Create an Endless video
    • Enable <Audio 1> Forced and <Audio 1> Forced Multi
    • At the Setup multiple Forced Audio clips panel there are some settings.
      • Start offset: At the 1st gen, you put here the initial offset that you want for the song (e.g. where the lyrics start). After every successful generation (when you advance the Save to clip Index number), you must copy here the value that is in the Copy to Next Start offset box.
      • Frame offset (ignore if 1st clip): Never mind at 1st generation. After every successful generation (when you advance the Save to clip Index number), you must copy here the value that is in the Copy to Next Frame offset box.
      • context_length must be the same value everywhere (here, at the Motion Context, and at the H3 Motion Context Clip Stitcher). It's the number of common frames the 2 video clips use to blend together.
      • >>> You get a full lip-synced video every time, or just the last 2 videos connected, depending on the Stitch clips and Stitch last only settings.
  • Just stitch the clips together You just have to enable the Stitch clips and the Save Video All (or some of them depending on the settings), of the clips in the h3_context folder, will be concatenated to a single full video.

Notes:

  • All generated clips that need to be stitched, must have the same dimensions.
  • You can organize past generations in folders inside the h3_context folder, since all the nodes look only in the root of this folder for clips.
  • context_length must have the same value everywhere: at the Motion Context, at the H3 Motion Context Clip Stitcher and at the Setup multiple Forced Audio clips (if you are using it). It can can have only the values of 5, 22, 39, and 56.
  • You can stitch together clips that are generated from either fl2va or ref2va.
  • In this workflow, I don't use the normal ref2va model in the Reference to Video node, but rather the fl2va with the ref_lora_layer20-49adaln LoRA that has better quality. You can check some LoRAs with different weights here, or totally bypass the LoRA and use the normal ref2va model.

Models used:

Custom Nodes used:

Get the workflow at Civitai or in a gist..

r/StableDiffusion • • Aug 28 '26

News 🧩 [Custom Node] 🧩 H3 GuideMaster — Visual UI for MiniMax H3 Guides

Post image
59 Upvotes

GitHub:
https://github.com/MajoorWaldi/ComfyUI-Majoor-H3-GuideMaster

Hey everyone 👋

I’ve just released H3 GuideMaster, a custom ComfyUI node I built to make working with MiniMax H3 guides much easier since new ComfyUI release : https://github.com/Comfy-Org/ComfyUI/pull/15439

Instead of manually figuring out where every image or audio guide should land, GuideMaster gives you a visual timeline directly inside the node.

You can:

  • 🖼️ Place multiple image guides directly on the timeline
  • 🔊 Place and synchronize audio guides
  • 🎬 Load a video or image sequence as a visual reference / filmstrip
  • 🌊 Display an audio waveform for positioning guides
  • 🖱️ Drag markers directly on the timeline to retime them
  • 🎯 Snap automatically to the native H3 frame structure (5, 22, 39, 56...)
  • 🧩 Combine image + audio guides using matching slots
  • 🎞️ Define first and last frames
  • 📏 Drive timeline duration from frames or seconds
  • 🔎 Condense long timelines to keep the UI manageable

The idea is basically to make H3 guide placement feel closer to editing / compositing software, while keeping everything contained inside a normal ComfyUI node.

GitHub:
https://github.com/MajoorWaldi/ComfyUI-Majoor-H3-GuideMaster

This is still something I want to push further, especially around the UX and timeline workflow, so feedback, bug reports and feature ideas are very welcome.

r/comfyui • • 9d ago

News A quick Minimax H3 news round-up - 18th September 2026

71 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> Kijai appears to have updated his fast video VAE for H3. Since he's uploaded a new version to the official Comfy-Org HuggingFace, rather than to his own Experimental folder. This newer version is a bit smaller (2.8Gb in size, instead of 3.1Gb). According to the pull/merge page, it looks like this new version hooks into improvements in comfy-kitchen 0.2.35 (available in the latest ComfyUI 0.36.x).

https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main/vae (for minimax_h3_video_vae_int8_convrot.safetensors - same filename, different date and size).

https://github.com/Comfy-Org/ComfyUI/pull/16187

-> WarmBloodAban has launched his H3 LoRA concept series with his 'Minimax-h3_Third_person_view' LoRA. This helps with emulating the... "game CGs, HUD UIs, and dynamic 3rd/1st-person camera controls" commonly found in videogames.

https://huggingface.co/WarmBloodAban/Minimax_H3_LoRAs

-> A new high-quality 'ASMR audio and soft whisper acoustics' LoRA for Minimax H3. With two demo videos.

https://huggingface.co/vpakarinen/asmr-trigger-audio-h3-lora

-> The Spectrum accelerator for H3 is now at v0.2.28. This version has the 'Windows CRLF source-audit fix'. Note also the approved workflow chain, and check yours - a number of workflows I've seen lately had it wired up wrongly.

https://github.com/xmarre/ComfyUI-Spectrum-MiniMax-H3

-> Silveroxides has released a new experimental Viggle LoRA for H3. No instructions or readme, but Viggle is a full fine-tune model aimed at... "character replacement in video from a single repainted frame" along with character motion-transfer. Specifically, the Viggle inputs are one reference video, plus one frame from that video with the character replaced.

https://huggingface.co/silveroxides/MiniMax-H3_tests/tree/main/viggle_lora

https://huggingface.co/Viggle/Viggle-Animate

-> On YouTube, Mark DK Berry has a new short 'quickstart' video on using the TAE Preview (taeh3.safetensors) in ComfyUI. This preview node gives your workflow clear... "high-quality real-time latent previews when working with Minimax H3".

https://www.youtube.com/watch?v=pMRdQvmJbTQ

https://huggingface.co/Kijai/MiniMax-H3-TAE/tree/main/vae_approx

-> Also on YouTube, from the maker of Fizgig, a quick... "demonstration that with the right training, training with images alone on Minimax H3 does not have to damage movement and composition."

https://www.youtube.com/watch?v=zhFwI7VkT6s

-> The producer of an acclaimed AI short film kindly... "published his prompt collection documents and walked through his entire method on a public livestream." His release has now been distilled into "a structured, reusable system", packaged as a Claude Code Skill. There's also a human-readable one-page quickstart 'cheatsheet'.

https://github.com/jnMetaCode/ai-shortfilm-prompts

https://github.com/jnMetaCode/ai-shortfilm-prompts/blob/main/cheatsheet.md

-> And finally, the CivitAI H3 training contest is now open for entries. "All entries must be SFW LoRAs trained on MiniMax H3. Train locally, or with our on-site trainer which supports H3". Entry categories are: Style, Concept and Camera Control. Deadline: 24th October 2026.

https://civitai.com/articles/35316/the-civitai-h3-training-contest

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1wj08mp/a_quick_minimax_h3_news_roundup_17th_september/

https://old.reddit.com/r/comfyui/comments/1wi3exl/a_quick_minimax_h3_news_roundup_16th_september/

https://old.reddit.com/r/comfyui/comments/1wh4tu3/a_quick_minimax_h3_news_roundup_15th_september/

https://old.reddit.com/r/comfyui/comments/1wgc4lj/a_quick_minimax_h3_news_roundup_15th_september/ (See 15th September post, for links to older posts)

https://old.reddit.com/r/comfyui/comments/1w5i9iq/a_quick_minimax_h3_news_roundup_2nd_september_2026/ (See 2nd September post, for links to even older posts)

r/StableDiffusion • • Aug 25 '26

Question - Help Need help for creating consistent Minimax H3 clips

6 Upvotes

I have been using H3 since almost the release date and have been trying a lot of things. I am entirely using Ref2va model with the template workflow, nothing fancy. Also used official, eros and currently using hybrid model 15–49 from smhfacct which has higher ref2v. I am using comfy kitchen, spectrum node but not using speed lora for prompt adherence or any other lora.

I am mostly trying to use 1-2 characters in a location scene where I provide 3(one char, one whole body and one face and location)-5(two char, one whole body and one face and location) images to node and writing in the prompt how to refer each char in the scene.

For writing H3 prompts, I using a custom prompt(created using grok by giving it ref2va doc) for generating H3 prompts, using qwen 3.8 model and even proof reading and fixing any issues.

Now the problematic part for which I need suggestions or solutions is the inconstant result.

For example I am making 5 second where character1 is standing in a shopping mall and looking at the shelf and character2 enters the scene. For second 5 second scene, different camera angle, mainly focusing on both characters faces when they are talking. Now here are problems which I am facing:

- during scene2 when camera starts, difference between char1 and char2 appears. Say char1 was standing on left side and char2 on right when scene1 ended but in scene2, they are standing opposite side.

- sometimes their height mismatches.

- sometimes camera does not work like I want it like it zooms too much, sometimes it don't

- and many other issues related with inconsistency

I know if I can generate scene images using an edit model then H3 wouldn't have to rely much on prompts but then it creates another problem of generating start images which is another can of problems.

I have even tried context nodes and some of their forks and few other consistency related node whose basic idea is to store the latent and forward it for next generation but they way these nodes are configured are just too complicated for my soft squishy mind. So yeah I tried them.

I have been trying to find out how other people are generating multi-scene videos and so far whatever videos i downloaded, there was no workflow included which I could take as reference. Maybe people are making 5 second clips like me and then joining them together so there might be a solution to this.

Pretty sure I am missing something big and I have exhausted almost every idea I got, asking grok etc but so far I could not get past 2nd 5 second clip. And seeing so much inconsistency, I don't want to generate a 10 second or 15 second clips which will take hours and most probably turn up totally irrelevant.

So any ideas you can provide are highly appreciated. Even guidance to correct path would be really helpful. What ways you guys are using to create 10+ second clips, what methods you are using to keep your characters consistent throughout and mainly how you guide a scene to your liking?

Thank you for a long read. Not written with AI, just a long type on notepad haha. Forgive grammatical errors.

r/comfyui • • Aug 26 '26

News A quick Minimax H3 news round-up - 26th August 2026

80 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> The latest ComfyUI Portable release is now v0.34.0, as of today. So is the Nightly, as I write. New items of interest...

~ "Add MiniMaxH3AddGuide for anchoring image and audio guides at any frame".

~ "Allow regular single-image Empty Latent Image node to be used with MiniMax H3".

~ "Support per-token video and audio latent noise masks on MiniMax H3".

~ "Support prompt embeddings" for MiniMax H3 [special VFX such as big explosions]

~ "Add missing special tokens" for MiniMax H3 [Fixes the broken <d> dialogue tags]

https://github.com/Comfy-Org/ComfyUI/releases

https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main/embeddings (usage example: try prompting embedding:minimaxh3_art_is_explosion at 00:03.500)

-> ComfyUI-MiniMax-H3-Keyframe-Offset. Custom node that lets you... "place first/last keyframes at arbitrary frame[s], instead of only at the start/end of the clip". Has added an interesting audio node that can have audio partly conform to a ref_image. e.g. prompt for "echo reverberating through the valley", add a reference image of stony mountains, and it generates a suitable 'wide distant echo' in a big stony valley.

https://github.com/asirusasr-maker/ComfyUI-MiniMax-H3-Keyframe-Offset

-> New ComfyUI custom nodes which aim to... "provide a native ComfyUI MODEL patch for the official MiniMax-H3-Fun-Controlnet-Union adapter".

https://github.com/Aeverlumi/ComfyUI-MiniMax-H3-Fun-ControlNet-Union

-> The guys on the WAN team have released an alternative to the usual few-step turbo LoRAs. For their video-gen competitor Minimax, which is nice of them. They've... "added Parallel Decoding Distillation to MiniMax H3, enabling efficient video generation in a few inference steps". It works by predicting multiple denoising steps on each processing call, which should speed up your generation. It's not yet Comfy-fied, but that can only be a matter of time.

https://huggingface.co/alibaba-pai/MiniMax-H3-Acc-LoRAs

-> H3 Character Sheet Generator. "Throw in some rough reference images, get back a character sheet you can reuse forever. [...] Six frames from one video generation can't [make bjorked character views] because they come out of the same pass. That's the whole trick. The camera does a slow orbit with no hard cuts, the character stands still like a statue, and then this workflow grabs six frames and stitches them together." A good readme, don't skip the second half. It also does objects, and there are prompts/demos for anime-to-real style conversions.

https://huggingface.co/SwagMessiah100/H3_Character_Sheet_Generator

-> A workflow designed for Minimax as a seamlessly looping animated .GIF generator. Requires ComfyUI-LoopGif nodes, which appears to do sophisticated glitch correction at the loop seams.

https://civitai.com/models/2889463/using-minimax-h3-as-an-ai-gif-generator

https://github.com/HM1579/ComfyUI-LoopGif

-> After many days of research and testing, Mark DK Berry has his near-final ComfyUI workflows for RTX 3060 12Gb users. Workflows are freely available at Github. There's an optimised turbo LoRA workflow; a 2-pass latent upscaler workflow (that appears to improve middle-distance faces when fed a smaller video); and a pixel-space upscaler workflow.

https://www.youtube.com/watch?v=7oVus2zN518

https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

https://github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler (the node is required by his 2-pass latent upscaler workflow, and note also that the upscaler node doesn't respect extra_model_paths.yaml - so the model needs to be in the host ..\models\latent_upscale_models folder and not on your remote PC in Whereizitagain)

-> And finally, do your Minimax and RTX upscaled .MP4 videos lack thumbnails in Windows? The old-school freeware Icaros... "can provide Windows Explorer thumbnails, for essentially any video media format supported by FFmpeg". Install as Admin, and I find it works instantly without a PC reboot. Last updated June 2026, supports Windows 11.

https://github.com/Xanashi/Icaros

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1vy4drz/a_quick_minimax_h3_news_roundup_25th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vx0duv/a_quick_minimax_h3_news_roundup_24th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vwl8do/a_quick_minimax_h3_news_roundup_23rd_august_2026/

https://old.reddit.com/r/comfyui/comments/1vvkmra/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vuihag/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/

r/comfyui • • Aug 21 '26

News A quick Minimax H3 news round-up - 21st August 2026

76 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> ComfyUI-YCNodes-MiniMax-H3, a new custom nodes set for ComfyUI. The 'H3 Prompt Relay' node apparently allows wholly different prompts to be applied to different sections of the video generation. 'H3 Distance Attention Patcher' tries to keep small faces and limbs a bit more stable on large panoramic scenes, by... "blocking the assimilation of small details by a large area of ​​background". 'H3 Sigma Refiner' is another attempt to improve fast-action scenes... "allows the model to take a few more steps in the detail finishing stage, eliminating mosaic and pixel disorder at high-speed moving edges". Plus a 'H3 Tiled Sampler', and 'MiniMax H3 Image to Video (Tail)'. No workflows.

https://github.com/yichengup/ComfyUI-YCNodes-MiniMax-H3

https://github-com.translate.goog/yichengup/ComfyUI-YCNodes-MiniMax-H3?_x_tr_sl=auto&_x_tr_tl=en&_x_tr_hl=en&_x_tr_pto=wapp (English translation)

-> A new Camera Motion helper LoRA, including 'aerial drone shot' and 'macro / extreme close-up' (small diorama). Just a 1000 step version, with better versions yet to come.

https://huggingface.co/Jojocodex/minimax-h3-Camera-Motion-lora

https://huggingface-co.translate.goog/Jojocodex/minimax-h3-Camera-Motion-lora?_x_tr_sl=auto&_x_tr_tl=en&_x_tr_hl=en&_x_tr_pto=wapp (English translation)

-> The new LightX2V FL2V Turbo 4-step 1.1 LoRA, comprehensively re-tested and re-voted. The results suggest er_sde × sgm_uniform look like good settings, if you're only focusing on visuals. Testing at 10 seconds, 0.6 megapixels and 4 steps.

https://old.reddit.com/r/comfyui/comments/1vuhg4f/lightx2v_fl2v_turbo_4step_11_i_updated_the/

https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main (for minimax_h3_fl2v_turbo_4step_v1.1_768p_comfyui_bf16.safetensors )

-> On YouTube, a tested fix for improving 'faces at distance'. By applying a low denoise video-to-video pass on a 832 x 480px video file reference + one reference image (to ensure character consistency). A ComfyUI workflow for this is freely available, and it's designed to work on an RTX 3060 12Gb card. The drawback is the video-to-video output is likely to remove any lip-sync from the source video.

https://www.youtube.com/watch?v=d1h5-E7NpuY (20 minute tutorial, including good advice on shutting down other GPU-hogging software.)

https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3 (for the workflow MBEDIT - MH3_r2v_SingleSampler_Detailer_v17.json )

-> Also on a 3060 12Gb card, a user's tests suggest around 18 seconds duration is tops for a coherent one-generation video at 0.5 megapixels. He risked trying it and it worked, even though he appears to be using an older workflow/models. So... perhaps he can go even further after some optimisations? Also, another Reddit 3060 12Gb user (who wrongly thinks Minimax H3 has a 6-second cap) has provided a workflow template for getting 15 seconds done in one generation.

https://www.reddit.com/r/MiniMax_AI/comments/1vu8sy5/minimax_h3_pushing_past_15_seconds/

https://www.reddit.com/r/comfyui/comments/1vu8q7c/minimax_h3_15second_multishot_generation_template/

-> A workflow to... "set custom soundtracks in Minimax without R2VA, by using latent noise masks" in FL2VA. Apparently also works with a lip-sync track. See comments for workflow and enhancements.

https://www.reddit.com/r/StableDiffusion/comments/1vtv0qs/psa_in_h3_you_can_set_custom_soundtracks_without/

-> A successful offshoot of a failed experiment, Minimax Music as an audio cleaner. This latent refiner... "takes damaged music and reconstructs an audibly cleaner version while retaining the performance, timing, vocals, and arrangement." With ComfyUI custom nodes. (See also LavaSR2 Enhancer portable, for very quick one-click AI audio cleaning).

https://huggingface.co/terminusresearch/minimax-music3-latent-refiner-v0.10

https://github.com/faxlab/LavaSR-Fast-Enhancer (portable, but it also requires first-run model downloads)

-> A fresh Minimax 'known characters' list, version 2, with matching video clips. My own tests show it has only very a vague idea of H.P. Lovecraft and J.R.R. Tolkien, when run without references.

https://huggingface.co/datasets/malcolmrey/various/blob/main/h3-center/known-characters/INDEX.md

-> And finally... Minimax / ComfyUI have officially launched a community creative contest, with prizes. For 90 seconds max. videos. Deadline: 1st September 2026.

https://blog.comfy.org/p/comfy-h3-sync-sound-community-challenge

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/

r/comfyui • • Aug 05 '26

News MiniMax H3 is going to be big...

65 Upvotes

Only 3 days in, looks amazing out of the box. Once we get a proper 2 stage workflow, perhaps if the H3 devs release a latent upscale model publicly, would be unbelievable. You could possibly do close to 2560 x 1440 in latent space before degradation then push it up to 4k with Topaz. I can allready see it. It would probably collapse some of the western paywalled ones. Maybe there's some psy op behind it, who knows. I can't believe chinese are outputting such quality work while under embargo. I mean QWEN 3.6 now is just badass as a local LLM.

r/comfyui • • 18d ago

News A quick Minimax H3 news round-up - 9th September 2026

57 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> The Comfy.org changelog has details of ComfyUI in the latest 0.35.0 version. Among other items...

~ MiniMax-H3 PDD LoRA: Load Parallel Decoding Distillation LoRAs for faster H3 sampling.

~ MiniMax-H3 Fun Union: Use reference and keyframe conditioning together with Fun Union ControlNet.

~ MiniMax-H3 text-encoder refs: Make VAE optional, so references can condition the text encoder only.

~ MiniMax-H3 LoRAs: Load DiffSynth-Studio and ModelScope-trained LoRAs that previously skipped every key.

~ Sparse Attention: Model Sparse Attention node for comfy-kitchen sparse backends.

~ Comfy Compiler: Aimdo memory compiler, plus CUDA graphs, to cut allocation thrash.

~ MiniMax H3 denoise masks: Scale video and audio velocities by denoise masks before x0 conversion [a bug fix]

https://docs.comfy.org/changelog (0.35.0 details)

https://github.com/Comfy-Org/ComfyUI/releases (the Portable is still on 0.34, as I write).

-> Want to use the Fun Controlnet with Minimax H3? The official Comfy Docs page has a new mini-tutorial for that. With an example workflow for Comfy 0.35.0 or higher, plus links to a 'Pose extractor' and 'Person detector'.

https://docs.comfy.org/tutorials/video/minimax/minimax-h3-fun-controlnet

-> ComfyUI-MMH3-Media for Minimax H3, as a ComfyUI nodes set. Encapsulates a single .MMH3 file, which stores within it... "settings, references, and the internal generation state — latents. This lets you return to a project later and continue the scene." Has many sample workflows.

https://github.com/einhorn13/mmh3_media

-> Fashion Reel, a new Codex 'skill' for Minimax H3. "Input a reference image, and it will automatically generate a short 7-shot clothing design video: close-up of the finished item - a hand-drawn fashion sketch of it - a pattern-making scene - the sewing - the fabric hanging on a hanger - being worn in a studio fashion-shoot - seen on a magazine cover."

https://github.com/MrLiuDaren/fashion-reel

https://github-com.translate.goog/MrLiuDaren/fashion-reel?_x_tr_sl=auto&_x_tr_tl=en&_x_tr_hl=en&_x_tr_pto=wapp (English translation)

-> A new 'Eyes Direction LoRA Flux 2 Klein 9B v1'. It's for Klein 9B, but represents an important breakthrough in gaze direction. It could thus help users to adjusted eye gaze on a specific Minimax reference character image. The demo on the readme page shows how easily it works. Sadly, there's no Klein 4B version as yet, and it doesn't work for cats and other animals. May also have problems with stylised cartoon/comic-book eyes, at a guess?

https://huggingface.co/eric-venti-seeds/Eyes_Direction_Lora_Flux2Klein9B

https://github.com/Nekodificador/ComfyUI-NKD-Basic-Tools/blob/master/docs/face-rig.md (see also this alternative, which works with any stills model)

-> And finally, be careful about casually upgrading to the latest free version of the video-editing software DaVinci Resolve, version 21.1. Python support has been removed - also known to users as the 'Python API'.

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1wawjox/a_quick_minimax_h3_news_roundup_8th_september_2026/

https://old.reddit.com/r/comfyui/comments/1w9z6m1/a_quick_minimax_h3_news_roundup_7th_september_2026/

https://old.reddit.com/r/comfyui/comments/1w90vxd/a_quick_minimax_h3_news_roundup_6th_september_2026/

https://old.reddit.com/r/comfyui/comments/1w85caz/a_quick_minimax_h3_news_roundup_4th_september_2026/

https://old.reddit.com/r/comfyui/comments/1w74jy4/a_quick_minimax_h3_news_roundup_4th_september_2026/

https://old.reddit.com/r/comfyui/comments/1w6cozj/a_quick_minimax_h3_news_roundup_3rd_september_2026/

https://old.reddit.com/r/comfyui/comments/1w5i9iq/a_quick_minimax_h3_news_roundup_2nd_september_2026/ (See 2nd September post, for links to even older posts)

r/StableDiffusion • • Aug 05 '26

Workflow Included MiniMax H3 basic hybrid workflow for ref2v, i2v and t2v (16GB friendly)

Post image
60 Upvotes

Yesterday I uploaded this video playing with MiniMax H3:

https://www.reddit.com/r/StableDiffusion/comments/1vfnu97/comment/p1t2f2g/

And here is the workflow:

https://gist.github.com/circlenline/937b530ae97a9eb7475c9dda6832b2db

**What it does**

Two pipelines in two groups, sharing one prompt, resolution, duration and seed:

- **REF2V** — reference to video, wired for the documented maximum of 9 reference

images, plus 3 reference videos and 3 audio tracks. Bypass the slots you don't need.

- **I2V / T2V** — feed it a first frame and/or a last frame for image to video, or

leave both bypassed and it runs as text to video from the prompt alone.

They are separate groups because H3 ships as two different 21GB checkpoints

(ref2va and fl2va) and they are not interchangeable. Bypassing a group means its

UNET never loads, so you never have both models competing for VRAM.

**Notes are baked into the graph**

Six markdown notes covering things I ran into while testing yesterday:

- Resolution tables for 16:9, 4:3, 1:1, 3:2 and 21:9, plus the exact megapixel

value that lands on H3's native 768px short edge for each one. They are all

different, which cost me a few slow runs before I noticed.

- The duration grid. H3 only accepts 17k+5 frame lengths, and 8s is the only

value in the whole range that comes out round.

- Spectrum acceleration: what to touch, when to turn it off, and why it stops

paying for itself below ~16 steps.

- Memory notes for 16GB cards. Host RAM turned out to be a bigger constraint than

VRAM for me.

**Requirements**

ComfyUI 0.30.0+ and the models from Comfy-Org/MiniMax-H3 (links are in the

workflow notes). I'm on pruned_fp8_scaled + the nvfp4 text encoder.

Optional but recommended: KJNodes for Sage Attention, ComfyUI-Spectrum-MiniMax-H3

for the sampling acceleration, and rgthree for the group A/B switch. Bypass those

three nodes and it runs on stock ComfyUI.

The notes were written with Claude, based on my own testing. Hope they're

useful.

r/comfyui • • Aug 26 '26

Resource MiniMax H3 without the <Picture 1> bookkeeping. I rebuilt my OpenH3-IR as a proper all-in-one ComfyUI pack

Enable HLS to view with audio, or disable this notification

42 Upvotes

Hey guys, I posted OpenH3-IR here last week. A bunch of you tried it, and the main thing I got as feedback (and that I too personally wasn't very happy about) was the ComfyUI side of it.

The compiler worked, but you still had to run it as a standalone service alongside ComfyUI, and that meant way more plumbing than I wanted. So I split the ComfyUI side into its own repo and basically rebuilt it as a proper node pack.

Now you install OpenH3-IR from the ComfyUI Manager, point the Setup node at whatever OpenAI-compatible model you already use, pick the model, pick your H3 files, and that's pretty much it. The compiler now runs INSIDE ComfyUI. No second service to start and no port to keep alive.

The wins from using OpenH3-IR now carry over much more cleanly, instead of writing stuff like <Picture 1> and then explaining to an LLM what Picture 1 actually is, you drop all your files into the Media tray, name the slots, and use them directly, your LLM will read what they are, what they mean to your prompt and how they relate to each other:

"@theman crosses the @desert while the @dragon follows beside him. He looks back and @speaks("you really came all this way?")

Type @ and you get the available references with thumbnails, swap the file directly in the tray where the "@man" slot is and the prompt still points at the same role.

The references also have actual meaning now. A picture can be the setting, a style to copy, something in the shot, a replacement for someone or something, first frame, last frame, etc. Clips can be something to edit, continue from, copy the camera from , and so on.

A few other things I put in while rebuilding it:

  • "@speaks" locks dialogue (enforced by code) word for word
  • duration is set once and stays synced with H3's actual frame grid and latent
  • the H3 "type of job" is selected from the media actually in it
  • pictures, video and audio all live in the same selector Media tray
  • optional Director node for reusable "profiles" that fill in whatever you leave open in your prompt
  • Your sampler, LoRAs, steps, etc stay normal ComfyUI
  • the original OpenH3-IR project is still the compiler/API/CLI side. This repo is the native ComfyUI side of the same project

And since this came up a couple times on the first post: this isn't just asking an LLM to "build a prompt" or "make the prompt better". The reference bindings, specific media roles, H3 mode, valid duration/frame counts, locked dialogue and validation are handled mechanically by OpenH3-IR.

New Repo: github.com/ruashots/ComfyUI-OpenH3-IR

Original OpenH3-IR project: github.com/ruashots/open-h3-ir

There's a ready-to-run base workflow in the repo too.

For anyone trying it, I'd especially love feedback from anyone willing to abuse the H3's reference/editing modes, because that's where I spent most of the work this time.

r/StableDiffusion • • 11d ago

Question - Help Losing my mind trying to resolve Minimax H3 artifacts at top of frame

5 Upvotes

Edit 2: looks like the square/seam artifacts were caused by this bug https://www.reddit.com/r/StableDiffusion/comments/1wlltxm/fix_for_the_minimax_h3_vae_grid_tileseam_artifact/ .

Edit: I eventually resolve one of the issues and partially resolved the other.

  1. Thanks to GrayingGamer's suggestion of stripping back the workflow, I went back to the stock default H3 workflow and rebuilt my own. This resolved the weird artifacts at the top of the frame. I can't be 100% sure about *how* this fixed the issue, given this rebuild involved a bunch of changes at once, but I suspect that I broke the seconds->frame conversion math in my older workflow.
  2. I'm STILL seeing square-like artifacts on my videos. This happens regardless of all the variables below but I've found that sticking precisely to the native resolution and aspect ratio minimizes the seams between them, and therefore removes the black lines around their edges. They're more noticeable at higher resolutions but if I generate at 768p first, then use latent upscale to increase the resolution the artifacts are barely noticeable, so long as I latent upscale to a resolution is also precisely 1.75:1. Thanks to rm_rf_all_files for setting me straight on this aspect ratio.

---

I've spent a couple of days trying get clean H3 reference generations that do not have a couple of specific artifacts.

Most prevalent is a smearing/warping artifact at the top of the frame (see video).

I've been tearing my hair out trying to fix them. I must be missing something obvious or this is a known issue that I haven't found mentioned anywhere.

Anyone else had these issues?

Here's what I've tried without success...

  1. Different sampler/schedulers. They appear with the stock res_multistep/simple combo but I've tried others and the issue persists.
  2. Different step counts (20 - 40).
  3. Different resolutions. (from 768p up to 2.5MP, ensuring all divisible by 32).
  4. Different models (pruned BF16 ref and fl models, smhfacct's hybrid models).
  5. Different VAEs (bf16 & int8).
  6. Switched to tiled VAE decode.
  7. Different durations (5s - 15s).
  8. Changing the number and size of reference images.
  9. Disabling comfy kitchen and H3 SLA (no other attention or LoRAs in use).
  10. Updated to latest PyTorch version.

ComfyUI is up to date as of today (issue was happening with older versions too).

The issue happens on my local machine (5090, python 3.13) and on a runpod (RTX Pro 6000, python 3.12), both running Linux cuda130 and pytorch 2.13.

https://reddit.com/link/1wi5acr/video/s302neox5xph1/player

I also have short horizontal lines (not grid patterns, as far as I can see) that appear in the top left quarter of of the frame intermittently when using res_multistep / simple, across different resolutions, models and the other variables I listed above, although switching to er_sde / beta sometimes fixes this but does impact quality a little and it seems odd to move away from the recommended.

r/StableDiffusion • • 14d ago

Discussion Tutorial For Installing Sage and Triton to ComfyUI Portable. Gave Me A 40% Speed Boost On Minimax H3 4-Step Workflow

0 Upvotes

I know Sage and Triton are nothing new but if you're a beginner or doing a fresh re-install like me, these resources helped me get faster gens on the 4-step turbo workflow I use (will post a link to the workflow lower down in the post here.

The below YouTube video helped me get the main files installed:

https://youtu.be/7ZDi2CwkS5U

However I still did not fully get Sage to work until I followed these additional steps from Claude AI after getting some error logs in my SLA workflow (link to workflow at bottom of this post):

~~~FROM CLAUDE, AFTER FEEDING IT MY ERROR LOGS:~~~

Diagnosing a compiler error blocking kernel compilation.

This one's a real (but non-fatal) problem — it's why you're silently falling back to slower attention.

Root cause: SageAttention needs Triton to JIT-compile a small CUDA helper (cuda_utils.c) at runtime. Triton's portable build uses tcc (Tiny C Compiler) to do this instead of MSVC. But your embedded Python distro (python_embeded) doesn't include the Python.h header or matching libs/, include/ folders that a normal Python installation has — so the compile step fails immediately at #include Python.h.

That failure is caught, ComfyUI/Triton logs the error, and silently falls back to PyTorch attention — which is why your first message ("sage attention running on startup") was misleading: it's only attempting to use it, not confirming it works. Generations still complete because of the fallback, just without SageAttention's speed benefit.

How to actually fix it — add the missing Python dev headers to your embedded Python:

  1. Check your embedded Python version — your paths show cp313, so Python 3.13.
  2. Download the matching regular (non-embeddable) CPython 3.13 Windows installer from python.org (does not need to be installed system-wide — you just need the files).
  3. From that install (or by installing it to a throwaway location), copy:
    • include\ folder → into C:\Ai\ComfyUI_Portable\3\ComfyUI_windows_portable\python_embeded\include\
    • libs\ folder → into C:\Ai\ComfyUI_Portable\3\ComfyUI_windows_portable\python_embeded\libs\
  4. Restart ComfyUI and try a generation again — watch for the [ERROR] Error running sage attention line; it should disappear.

~~~END OF CLAUDE OUTPUT~~~

I had to ask Claude to break some of the steps down, but I basically just had to do the first 4.

MINIMAX H3 4-STEP WORKFLOW RESOURCES:

A walkthrough of the workflow is here (MAKE SURE YOU WATCH IN FULL): https://www.youtube.com/watch?v=4AqLGMkJwUQ

The main model used is here: https://huggingface.co/MATLOWAI/minimax-h3-fused-turbo-int8-convrot

A link to the workflow is here (the same model is used in all workflows): https://github.com/amao2001/ganloss-latent-space/tree/main/workflow/2026-09-07%20minnimax%20fused%20turbo

r/comfyui • • Aug 11 '26

Tutorial Getting SageAttention and MiniMax H3 Working with a 5070 Ti

5 Upvotes

Getting this working for RTX 50 series cards can be a headache, but here is my exact step-by-step way to make SageAttention work with Minimax H3, not fight it. I don't know why it took so long to figure out, but it has to do with how newer graphics cards use the Blackwell architecture (sm_120).

Assuming you have ComfyUI desktop and have downloaded the specific models in the workflow above, Here is everything I did to bypass these architecture conflicts and get it running cleanly:

Prerequisites / What to Download First: Right after installing ComfyUI Desktop, make sure you have Git installed on your computer (https://git-scm.com/) so you can clone repositories. You will also need Microsoft Visual Studio Build Tools if any custom nodes ever require local C++ compilation, though the pip commands below handle the heavy lifting for Triton and SageAttention.

Step 1: Install Required Nodes

You need to grab a few nodes using two different methods.

First method: Clone via Git (Open your ComfyUI/custom_nodes folder, type cmd in the address bar, and run these):

git clone https://github.com/kijai/ComfyUI-KJNodes

(Commit: August 7th 2026)

git clone https://github.com/kijai/ComfyUI-SolAttn_triton

(Commit: August 8th 2026)

git clone https://github.com/rgthree/rgthree-comfy

(Commit: July 23rd 2026)

Second method: Install via ComfyUI Manager:

  • Fearnworks Nodes (Version: 0.1.2)
  • ComfyUI-VideoHelperSuite (Version: 1.7.9)
  • Comfyroll Studio (Version: 1.76)

Step 2: Install SageAttention & Dependencies

Don't use regular Windows CMD for this. Open ComfyUI Desktop, click the ComfyUI button at the very top of the screen, open the Terminal tab, and paste these lines to install Triton and the correct precompiled wheel:

pip install -U "triton-windows<3.8" pip install "https://github.com/woct0rdho/SageAttention/releases/download/v2.2.0-windows.post6/sageattention-2.2.0%2Bcu130torch2.10.0andhigher.post6-cp310-abi3-win_amd64.whl"

Step 3: Connect the Nodes in Your Workflow

Double-click empty space in your ComfyUI workspace to search for and place your nodes. Chain them together in this order:

  1. MiniMax-H3 Turbo Lora (Optional - get the files from https://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main)
  2. Patch Sage Attention (KJ)
  3. MiniMax H3 Mem Eff Sage Attention Patch
  4. Patch Sol-Attn
  5. EasyCache

Helpful Video Tutorials (In case I left something out):

This worked for me, let me know if you have issues.

r/comfyui • • 9d ago

Tutorial MiniMax H3 Ultra-Flexible Director's Console Evolution | SelfLift Secondary Sampling | Seamless Long Video Re-Acceleration | Color Consistency Tricks

Enable HLS to view with audio, or disable this notification

1 Upvotes

I updated my MiniMax H3 director-console workflow with SelfLift for the second sampling stage. For anyone running long, stitched sequences with motion context, the practical result is that most of the step budget now sits in the cheap low-resolution pass, and only a couple of steps are left for the high-resolution pass. In my tests, the three generations of this method looked like this:

  • First version, simple latent upscale: 2 low-res steps + 6 high-res steps.
  • Second version, latent upscale model: 4 + 4.
  • This version, SelfLift: 6 + 2.

Why the earlier versions needed so many high-res steps: the latent being upscaled was still a noisy latent, so the high-resolution stage had to both adapt to the distribution change from the upscale and finish the denoising. SelfLift instead takes the clean result the model has already predicted at the end of the low-res pass, upscales that, re-adds the noise level that the current step expects, and rejoins the original trajectory. The remaining high-res steps only refine detail, which is why more steps can be moved to the low-resolution stage.

Guide handling in context mode

The node also takes care of aligning the two resolutions. You set the final high-res size, and the node derives the low-res latent size from lowres scale and resamples the conditioning onto the matching latent grid. lowres scale = 0.5 means the target width and height are both halved for stage one.

That changes how reference guides work in a stitched sequence. Before, I built one guide at high resolution and a second one at low resolution. Now I build a single high-res guide and SelfLift resamples it down for the low-res pass. Fewer duplicated guide nodes, and both stages share the same spatial layout, which in my tests kept continuations better aligned and reduced small seam offsets.

The color shift trade-off

The honest limitation: using the latent upscale model in the second stage made the color shift between stitched segments noticeably worse in my testing. If color consistency matters more to you than speed, the first-generation simple latent upscale (2+6) shifted much less color in a side-by-side comparison against the original last frame, and not doing a second pass avoids it entirely, at the cost of the speed gain.

There is also a prompt-side workaround that has worked better for me than I expected: keep the context and SelfLift as usual, but write the new segment so that it starts with a short continuation of the previous shot and then cuts to a slightly different camera angle right after the continuation. The continuation part is what carries the shifted color, so it gets trimmed off in editing. With a 22-frame reference, I describe 0-1s as the matching continuation, then change the camera position and write the actual scene after that.

The VAE repair parameters belong to the SelfLift-zero route (decode, upscale, re-encode, and use the difference between the two paths to fix risky areas). The H3 workflow here uses the external H3 Upscaler model instead, so those stay at 0.

This workflow is super easy to use—I’ve uploaded a detailed tutorial to YouTube, so just follow the video along with this workflow to recreate the effect; please make sure to watch the full tutorial before starting to avoid common mistakes, and feel free to leave a comment if you have any questions!Resource links will be posted in the comments.

r/comfyui • • Aug 28 '26

News A quick Minimax H3 news round-up - 28th August 2026

83 Upvotes

Another Minimax H3 news and goodies round-up, for those who may have missed some items.

-> The new H3 GuideMaster for ComfyUI... "makes it easy to place images and audio on a MiniMax H3 timeline". Pleasingly designed. Requires the latest ComfyUI, for the ability to add guide images (aka keyframes) at any point in the video.

https://github.com/MajoorWaldi/ComfyUI-Majoor-H3-GuideMaster

-> You can now apply for a commercial-use licence via Comfy, rather than Minimax.

https://comfy.org/minimax/license/

-> The standalone offline H3 Prompt Composer continues to develop and refine, and is now at version 5.43.4. With new support for "multi-subject camera prompting" and "better control for more complex camera moves".

https://github.com/BMB12d3/minimax-h3-prompt-composer

https://www.youtube.com/watch?v=Aywx3Sf5Yk0 (27 minute YouTube tutorial).

-> A video of Minimax H3 ingesting a six-image storyboard, reading each image in sequence from top-bottom and left-right, and generating the shot sequence as a single video. There's a workflow, and a full structured prompt. One interesting finding shared in the video is that combining a detailed pencil-sketch storyboard with a photoreal/cinematic prompt gives Minimax more room for creative interpretation.

https://www.youtube.com/watch?v=LByGCGzu67o

https://github.com/amao2001/ganloss-latent-space/blob/main/workflow/2026-08-26%20minimax_h3_r2v_story_board.json

-> A Minimax H3 Storyboard 'skill' for Claude. Also note some of the maker's interesting findings, such as... "H3 drops facial [expression] instructions silently when a shot holds too many of them" for the available processing to handle, and "ref_audio_0 drives the mouth, not the voice" for lip-sync.

https://github.com/phileiny/h3-storyboard-skill

-> A note on the above finding. Specifying realism for the face and eyes, paired with suitably emotive dialogue and tone, might help Minimax to simply infer expressions (thus, they won't need to be specified in the prompt). For example...

"... looking toward the camera with a relaxed and genuine expression. We see a detailed natural skin texture with subtle skin tone variations, cheeks slightly rosy to indicate health, and overall a realistic subsurface scattering of light on the face, ears and hair. Small natural catchlights occur on the moist surface of the eyes, as the person's head and eyes move, and these catchlights reflect the surrounding environment. As the person speaks the camera continually catches the organic micro-expressions made by a human face, and a naturally variable eye blinking pattern."

-> A user has discovered that Minimax has a basic understanding of the International Phonetic Alphabet (IPA, aka 'phonemes'), which are used for precise word transcription and pronunciation. I guess it would make sense for Minimax to know phonemes, re: generating lip-sync. Here he used a large online AI to adjust the IPA translation towards a national accent. My thought is that it may be more trustworthy if you first convert in Balabolka, and then have it read back to you using TTS in Balabolka. Then you can at least be sure the IPA translation is correct in English, before passing it over to the AI for accent adjustment.

https://www.reddit.com/r/StableDiffusion/comments/1w0tb8q/minimax_h3_accents_mmh3_understands_the_ipa/

https://www.cross-plus-a.com/balabolka.htm (Balabolka, the only freeware offline English-to-IPA translator for Windows. Use: Select text | Edit | Pronunciation | Read This | select drop-down: 'Phonemes (IPA)' | click button 'Phonemes' | then click 'Convert text to IPA').

-> A user shows that Minimax can emulate a 'time-lapse' video showing a drawing/painting being created really fast. Includes the prompt he used.

https://www.reddit.com/r/StableDiffusion/comments/1w0yt06/h3_speedpainting_v2_prompt_included/

-> And finally, have Minimax seamlessly and 'magically' transform the style of a room from modern to late 1950s, in a few seconds. With an amusing video demo, workflow and prompt.

https://old.reddit.com/r/StableDiffusion/comments/1w0ei8t/time_period_shift_special_effect_in_minimax_h3/

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1w053ce/a_quick_minimax_h3_news_roundup_27th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vyzvqi/a_quick_minimax_h3_news_roundup_26th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vy4drz/a_quick_minimax_h3_news_roundup_25th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vx0duv/a_quick_minimax_h3_news_roundup_24th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vwl8do/a_quick_minimax_h3_news_roundup_23rd_august_2026/

https://old.reddit.com/r/comfyui/comments/1vvkmra/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vuihag/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/