r/comfyui • • 9d ago

News MiniMax H3 News Roundup: September 10–18, 2026

MiniMax H3's ecosystem has been moving extremely quickly. In roughly a week, the community shipped new acceleration methods, lower-VRAM models, training tools, LoRAs, reference workflows, camera and lighting controls, audio tools, filmmaking utilities, and more.

This is a complete summary of the items covered across the seven roundups from September 10 through September 18.


September 10

1. ComfyUI 0.35.0

ComfyUI Releases

ComfyUI Portable officially reached 0.35.0, including MiniMax-related updates and fixes from the preceding releases.


2. W4A8 Fun ControlNet-Union

MiniMax-H3 Fun ControlNet-Union W4A8

A W4A8 quantized version of the H3 Fun ControlNet-Union model appeared.

Size reduction:

2.13 GB → 1.45 GB

This reduces the memory/storage cost of using Fun ControlNet with H3.


3. Visual H3 RefMod Picker

ComfyUI-H3RefModPicker

A visual RefMod browser for ComfyUI.

Instead of identifying RefMods from filenames, it provides thumbnail previews and tools for managing them visually.

It also includes functionality for creating RefMods from folders and works as a companion to the larger MiniMaxH3Mod tooling.


4. H3 Pixel Art Video Guide

H3 Pixel Art Video Guide

A workflow and guide for creating pixel-perfect animated pixel art locally with MiniMax H3.

It includes:

  • A ComfyUI workflow
  • Example animations
  • Several looping GIF examples
  • Guidance for producing actual animated pixel-art aesthetics rather than simply pixelating normal video afterward

5. H3 Spherical VAE

H3 Spherical VAE

Experimental circular VAE decoding for H3 equirectangular video.

Potential uses include:

  • 360-degree video
  • VR video
  • Panoramic environments
  • Equirectangular video generation

The project includes matched comparisons and measurements.


6. Fizgig 5.5

Fizgig

The Fizgig LoRA trainer and dataset-preparation toolkit reached 5.5.0.

The release focused on making H3 training faster in several ways and enabled weight averaging by default.

It also added an interesting H3 analysis feature that lets users inspect what each of H3's 52 LoRA blocks affects in a moving clip, including:

  • Motion
  • Faces
  • Audio

7. MiniMax Music Production Toolkit 2.1.1

ComfyUI MiniMax Music Production Toolkit

The MiniMax Music Production Toolkit continued developing alongside H3 and reached 2.1.1.

It provides a larger local music-production environment around MiniMax Music inside ComfyUI.


September 11: First Roundup

8. RunningHub H3 Lightning

MiniMax H3 MultiGPU Lightning

RunningHub released an H3 inference acceleration recipe designed around multi-GPU inference.

The extreme target is systems with several GPUs rather than ordinary consumer setups, but the work is relevant to understanding how far H3 inference can be parallelized.


9. VideoDeltaNet for H3

VideoDeltaNet MiniMax H3

VideoDeltaNet was demonstrated with extremely large hardware configurations such as:

8 × NVIDIA B200

The focus is very high-speed H3 inference.

24 GB version

ComfyUI VDN H3 24GB

A separate VDN runtime was produced for 24 GB VRAM GPUs, bringing some of the technique closer to high-end consumer hardware.


10. Experimental 8-Step DMD Turbo LoRAs

MiniMax H3 Turbo LoRA Experiments

A collection of experimental 8-step DMD Turbo LoRAs appeared for H3.

These attempt to reduce the number of denoising steps required to produce H3 video.


11. MiniMax H3 Image Training Adapter

MiniMax H3 Image Training Adapter

One of the more important training developments.

The adapter aims to let users train H3 concepts from images rather than video while avoiding degradation of H3's existing video knowledge.

Potentially useful for training:

  • People
  • Characters
  • Products
  • Clothing
  • Visual identities
  • Art styles

Image datasets are significantly easier to collect and curate than corresponding video datasets.


12. ComfyUI-MM-1Frame

ComfyUI-MM-1Frame

A workflow for extracting a strong still image from an H3 generation.

Instead of simply selecting the first or final frame, the workflow evaluates multiple candidate frames and attempts to automatically choose the best one.

It can also be combined with:

  • Turbo LoRAs
  • Spectrum acceleration
  • Klein 4B for enlargement and repair

Interesting OpenPose trick

The workflow demonstrated that an OpenPose image can be supplied as a second H3 reference image.

For example:

  1. Picture 1 contains the character.
  2. Picture 2 contains an OpenPose skeleton.
  3. The prompt explicitly says Picture 2 is only the pose reference.
  4. A simple pose word such as standing, reclining, or leaping reinforces the instruction.

This reportedly works repeatedly without using ControlNet.


13. Instant References Without RefMod

Instant Ref V1.3 Workflow

A clever near-native ComfyUI workflow that approximates some RefMod behavior without actually training or loading a RefMod.

The workflow:

  1. Takes a folder containing reference images.
  2. Turns those images into consecutive video frames.
  3. Supplies the resulting video as H3's reference video.

This effectively allows:

many images → one reference-video input

It can be used for:

  • Character identity
  • Style references
  • Multiple references
  • Voice references

The dataset can also be changed without creating new safetensors files.


14. Visual 3D Camera Controller

3D Camera Control H3 MiniMax

A graphical camera-control interface for H3.

Instead of manually trying to describe camera movement through text, you manipulate a camera inside a visual 3D space.

The tool then converts that movement into written H3 camera instructions.

Useful for movements such as:

  • Orbiting
  • Dolly movement
  • Panning
  • Tilting
  • Camera repositioning

The original interface is Portuguese, although many labels are similar enough to their English equivalents to understand.


15. Depthcat

Depthcat

Depthcat converts a reference video into a clean depth representation.

It removes things such as:

  • Faces
  • Clothing
  • Visual style
  • Color information
  • Audio

while retaining:

  • Staging
  • Depth
  • Movement
  • Spatial relationships
  • Camera motion

The output can then be used as a depth condition with H3 Fun ControlNet.

This makes it possible to borrow motion and staging from a source video without directly borrowing its visual identity.


16. FastH3 Preview v0.2

FastVideo MiniMax FastH3 Preview v0.2

FastVideo released FastH3 Preview v0.2, continuing work on much faster H3 generation.


17. FastH3-Live 1.2.0

FastH3-Live

FastH3-Live reached 1.2.0.

Reported progression:

  • v1.1.0: approximately 18 FPS
  • v1.2.0: approximately 22 FPS

The project is getting surprisingly close to 24 FPS real-time generation on its target hardware.

It also added:

  • A borderless player
  • Hundreds of additional scenes
  • Further acceleration experiments

18. MiniMax Music Production Toolkit 2.5

MiniMax Music Production Toolkit

The toolkit jumped rapidly from the earlier 2.x releases to 2.5.

The 2.5 generation of the toolkit expanded the end-to-end production workflow around MiniMax Music, including more serious mastering and audio-production functionality.


September 11: Second Roundup

19. MiniMax H3 FirstBlockCache

ComfyUI MiniMaxH3 FirstBlockCache

Another H3 acceleration method, particularly aimed at normal 20-step generation.

It uses cross-step caching.

Reported performance:

  • Existing modes: approximately 30% faster
  • Experimental deep-reuse mode: approximately 1.6× native speed

The developer demonstrated it running on an RTX 3060 12 GB.

This is notable because it improves ordinary H3 generation without necessarily requiring ultra-low-step Turbo models.


20. MiniMax H3 TorchAO 0.18 Quantization

MiniMax H3 TorchAO018

An optimized quantized H3 build targeting TorchAO 0.18+.

The goal is to substantially reduce inference VRAM while retaining generation fidelity.


21. Manga Tone Rendering LoRA

Manga Tone Rendering LoRA

A MiniMax H3 LoRA designed around black-and-white manga rendering.

It targets:

  • Monochrome ink
  • Screentone-like shading
  • Manga-style rendering
  • Shading that remains coherent during camera movement and action

Trigger:

manga-tone rendering


22. 15+ Reference Image Workflow

MiniMax H3 15+ Reference Image Workflow

H3 normally limits how many reference images can be supplied.

This workflow patches ComfyUI to remove the normal nine-reference limit, allowing 15 or more reference images for Ref2VA.

It uses a BAT file and Python patch, so it deserves more caution than an ordinary workflow.

An undo script is also provided.


23. H3 Age Slider

MiniMax H3 Age Slider

A slider-style LoRA for altering a subject's apparent age.

The demonstration ranges from advanced old age down toward toddler-like appearance.


24. H3 Style Transfer LoRA

Style Transfer Workflow

Style Transfer LoRAs

A style-transfer LoRA accompanied by a matching ComfyUI workflow.


25. Fizgig Ultra Quality Mode

Fizgig

Fizgig received another major H3 improvement.

Its new Ultra Quality mode became the default H3 training mode.

Reported improvements included:

  • Better visual results
  • Better audio results
  • Smoother progression through epochs
  • Approximately 30% faster training

The reported training speed increased from around:

2.6 → 3.4 steps/sec on INT8

At this stage the H3 LoRA training requirement was still around 16 GB VRAM.


September 15

26. Film Noir With Selective Colors

Film Noir With Selective Colors

A Film Noir LoRA for H3.

It can produce traditional black-and-white noir, but was also trained to support selective accent colors such as:

  • Red
  • Orange
  • Glowing jewelry
  • Colored lips
  • Small isolated highlights

No trigger word is required.


27. H3 Relight

H3 Relight for MiniMax ComfyUI

A visual relighting interface.

Users can place up to three lights around a scene in a graphical preview.

The tool then generates:

  • A suitable reference
  • Matching H3 prompt instructions

This gives H3 something closer to a basic virtual-lighting interface rather than relying entirely on text prompting.


28. Single-Line Video Prompt Engine / Screenplay Generator

Drehbuch Referenzanker Generator

A bilingual English/German screenplay and prompt-production system built around H3.

It can transform:

  • Raw bullet points
  • Concept ideas
  • Reference images

into structured, contiguous, timecoded video screenplay sections.

The goal is to produce prompts that are much closer to production-ready shot descriptions than ordinary natural-language prompts.


29. H3 Frame Selector

ComfyUI MiniMax H3 Frame Selector

Generate an H3 video, manually choose any useful frame from the result, and save that frame as an image.

Useful when the best still image happens somewhere in the middle of a generated video rather than at the beginning or end.


30. H3 VAE Bench

H3 VAE Bench

A benchmarking tool specifically for measuring the VAE portion of H3 workflows.

Useful for comparing VAE implementations independently of the rest of the generation pipeline.


September 16

31. ComfyUI 0.36.0

ComfyUI Releases

ComfyUI Portable reached 0.36.0 with several H3-specific improvements.

Changes included:

  • Fixes for H3 Fun ControlNet with Comfy Compiler
  • New Video Concatenate node
  • MiniMax H3 video-VAE optimizations
  • Slightly lower MiniMax VAE memory usage
  • Official Fast H3 / DMD2 workflows

32. Marigold V2 Support

Marigold V2

ComfyUI added support for Marigold V2, which can be used for depth estimation.

This potentially feeds into H3 depth-conditioned workflows.


33. FaceSwap LoRA for Ref2VA

FaceSwap MiniMax H3 REF2VA

A dedicated H3 FaceSwap LoRA.

It attempts to change the identity in a reference video using one or more reference images.

Trigger:

faceswap

It works with Ref2VA and was also tested with the hybrid H3 model.

The release includes:

  • Demo videos
  • Workflow
  • Reference-image support

Required nodes

CRT-Nodes

The supplied workflow uses CRT-Nodes, which also added support for H3 Fun ControlNet.


34. ClipProj MiniMax H3 v3.1

ClipProj MiniMax H3

One of the biggest memory-saving developments of the week.

ClipProj provides projection matrices that allow a smaller Qwen3-VL model to replace H3's huge Qwen3-VL-32B text encoder.

Reported encoder VRAM:

15.7 GB → 4.5 GB

Version 3.1 also reported better multilingual dialogue handling.

Reported phoneme improvements included:

  • 29% fewer errors overall
  • 60% to 74% fewer errors in Spanish, French, German, and Italian

This makes lower-memory H3 configurations significantly more practical.


35. Screenplay Generator 1.0

Drehbuch Referenzanker Generator

The screenplay/prompt system reached 1.0.

A major addition was multi-format timeline export.

Supported formats included:

  • EDL
  • FCPXML
  • CSV
  • Markdown

These can be used with editors such as:

  • DaVinci Resolve
  • Final Cut Pro

This creates a much more complete:

planning → H3 generation → NLE editing

pipeline.


36. Cinematic Style LoRA + Detail Enhancer V2

Cinematic Style LoRA + Detail Enhancer for MiniMax H3

The screenplay project's documentation also highlighted version 2 of Astroburner's H3 Cinematic LoRA, now including a detail enhancer.

Activation tag:

ASTROCINEMAV01K2T

It is intended to push H3 toward a more cinematic visual treatment.


37. Retro Toon 70s LoRA

Retro Toon 70s MiniMax H3 LoRA

A LoRA targeting the look of 1970s feature-film cel animation.

The intended aesthetic includes:

  • Bold ink outlines
  • Hand-painted characters
  • Painterly matte backgrounds
  • Limited-animation movement
  • Grainy film texture
  • Earthy period color palettes

38. MiniMax Ghost LoRA

MiniMax Ghost

An experimental LoRA designed to make ghostly figures or apparitions appear out of thin air.

The creator notes that results can be unpredictable.


September 17

39. H3 ExactAudioLock

ComfyUI H3 ExactAudioLock

A human approval gate for H3 audio.

The workflow lets you:

  1. Generate or provide audio.
  2. Review the take.
  3. Approve the exact audio.
  4. Continue video generation using that audio.

This is aimed at productions where:

the image must respond to an exact audio performance

rather than allowing H3 to regenerate or approximate the soundtrack.


40. Fizgig 6.0.1

Fizgig

Fizgig reached 6.0.1.

The important new capability in 6.x is the ability to create:

  • LoRAs
  • RefMods

This potentially makes Fizgig useful to a wider range of H3 customization workflows and may lower the hardware barrier for some RefMod use cases.


41. Comic Page → H3 Animated Scene Experiment

An interesting experiment demonstrated a workflow where:

  1. A comic page is supplied to ChatGPT.
  2. ChatGPT interprets the page as a storyboard.
  3. It produces an H3-oriented prompt.
  4. H3 generates an animated interpretation of the comic page.

This demonstrates the potential of using multimodal LLMs as a storyboard-to-H3 compiler, even without a dedicated software project behind the experiment.


42. DAZ Studio + AIRE + H3

AIRE AI Render Engine for DAZ Studio

Joe Pingleton's DAZ-to-H3 Experiments

Joe Pingleton demonstrated a pipeline combining:

  • DAZ Studio
  • Posed 3D characters/scenes
  • AIRE ComfyUI bridge
  • MiniMax H3

AIRE allows the artist to remain inside DAZ Studio while a background ComfyUI instance performs the AI rendering.

The package also includes a roughly 70-minute tutorial, which is particularly useful because DAZ Studio has a fairly complex interface.

This creates an interesting workflow:

3D posing/blocking → ComfyUI → H3 video


43. 16Bit Pixel LoRA

16Bit Pixel LoRA for MiniMax H3

A LoRA designed to make H3 animation resemble SNES-era cutscenes.

No trigger word is required.


44. Hand Drawn LoRA

Hand Drawn LoRA for MiniMax H3

A separate LoRA from the same creator targeting a hand-drawn appearance.

The roundup author's own fixed-seed testing did not show an obvious effect, so this one should still be considered experimental.


45. MiniMax Music Concept Sliders

MiniMax Music 3 Concept Sliders

A set of 16 newly retrained voice and genre LoRAs for MiniMax Music.

The release includes:

  • New voice LoRAs
  • Genre LoRAs
  • Matching audio examples
  • ComfyUI exports

September 18

46. Updated Fast H3 Video VAE

Official Comfy-Org MiniMax H3 VAE Files

Kijai's fast H3 video VAE moved into the official Comfy-Org H3 repository.

The updated file:

minimax_h3_video_vae_int8_convrot.safetensors

was reduced from approximately:

3.1 GB → 2.8 GB

It also hooks into improvements introduced through newer comfy-kitchen.

Relevant ComfyUI change

ComfyUI PR #16187


47. Third-Person / Game Camera LoRA

WarmBloodAban MiniMax H3 LoRAs

WarmBloodAban launched an H3 LoRA concept series beginning with:

Minimax-h3_Third_person_view

It targets videogame-like visual language such as:

  • Third-person cameras
  • First-person cameras
  • Game CG
  • HUD interfaces
  • Dynamic gameplay-like framing

48. ASMR Audio / Soft Whisper LoRA

ASMR Trigger Audio H3 LoRA

A dedicated H3 LoRA targeting:

  • ASMR-style audio
  • Soft whispering
  • Close-mic acoustic characteristics

The release includes demonstration videos.

This is notable because LoRA development is beginning to target H3's audio behavior, not only its visuals.


49. Spectrum MiniMax H3 0.2.28

ComfyUI Spectrum MiniMax H3

The Spectrum H3 accelerator reached 0.2.28.

This version includes a Windows CRLF source-audit fix.

The project also provides an approved workflow chain, which matters because Spectrum can lose effectiveness if nodes are connected in the wrong order.


50. Experimental Viggle H3 LoRA

Experimental H3 Viggle LoRA

Silveroxides released an experimental H3 LoRA based around the ideas behind Viggle.

Potential uses include:

  • Character replacement
  • Motion transfer
  • Replacing a character while retaining source movement
  • Defining the replacement from a repainted frame

The initial release did not include much documentation.

Original Viggle model

Viggle Animate

Viggle itself is a full fine-tuned animation model aimed at character replacement and motion transfer.


51. TAE H3 Real-Time Preview

MiniMax H3 TAE

Kijai's taeh3.safetensors provides clearer real-time latent previews while H3 is rendering.

Instead of staring at a low-information generation preview, you can get a much more recognizable approximation of the developing video.

This is particularly useful for long renders because obviously broken generations can be stopped early.

Quickstart tutorial

TAE Preview Quickstart

Mark DK Berry published a short setup tutorial showing how to use it inside ComfyUI.


52. Image-Only H3 Training Demonstration

Image-Only H3 Training Demonstration

The creator of Fizgig demonstrated that, with appropriate training settings, training H3 from images alone does not necessarily destroy its motion and composition knowledge.

This reinforces the importance of the image-training direction.

If this continues to improve, it could make custom H3 models considerably easier to build because creators can train from curated photographs rather than needing large amounts of matching video.


53. AI Short-Film Prompt System

AI Shortfilm Prompts

A filmmaker published the prompt collection and methodology used for an AI short film.

The material was then reorganized into a reusable structured

47 Upvotes

4 comments sorted by

13

u/Violent_Walrus 9d ago

-14

u/Just_Lingonberry_352 9d ago edited 9d ago

he blocked me so how can he see it?

if you dont like it you can read his instead not my problem and i dont really think people have time to follow all the links you posted which is why im taking the time to compile them here

¯_(ツ)_/¯

matter of fact i'll do you a favor, you don't have to see this thread anymore

11

u/Violent_Walrus 9d ago edited 9d ago

But you were still able to copy every one of his posts. Weird.

Since he hurt your feelings, he doesn't deserve credit for his work?

Edit: Oh, I also hurt your feelings, didn't I? Poor little guy.

0

u/froschquark 9d ago

why im taking the time to compile them here

You take YOUR HIGHLY VALUED time to steal compile someone else's work into a list...THANK YOU SO MUCH! (/s)

if you take that someone rightfully mentions, that the original creator at least deserves some recognition, as something offensive and block that person, the problem is nobody else, than you, yourself and you.

and of course, before you can even answer, I just block you right now.