I have been experimenting with MiniMax H3 FL2VA as a pseudo-image generator in ComfyUI.
This is not a native text-to-image mode. H3 generates a short sequence, the video VAE decodes it into an image batch, and Image From Batch extracts one frame as the final still.
Portraits and cinematic images worked well, but the biggest surprise was typography and graphic design. I tested magazine layouts, posters, dashboards, infographics, charts, fantasy key art, and phone-style photography.
The outputs are not perfect, but H3 follows detailed art direction surprisingly well. It understands typography hierarchy, layout structure, palettes, charts, icons, ornament, and the relationship between text and imagery. The results become much stronger when the prompt defines the complete design instead of asking for a generic poster or infographic.
There can still be spelling mistakes, fake microtext, inaccurate chart data, and video-VAE artifacts, so the results need inspection. Still, this looks very useful for posters, covers, key art, presentation visuals, design exploration, and infographic drafts.
Parameter note
In my setup, these values produced the best results:
INT Length: 8
Image From Batch Index: 8Length controls the short sequence generated by H3. After decoding, Batch Index selects which frame is saved as the image.
The 8/8 combination is based only on my experiments. Preview the full decoded batch and test nearby values, since the cleanest frame may vary depending on the prompt, resolution, checkpoint, and node implementation.
Description
FAQ
Comments (11)
cannot download the "config" file. And if I can download where in the comfyui structure do I put it? or is it going to have some key value for the encryption node?
its just a comfyui workflow. download that .json file > drag and drop inside the comfyui
@reverentelusarca why does the workflow require some encryption thing?
It could be something weird on my machine... I don't know why... can't get the nodes to work and when I try to update the nodes "used in the workflow" it has something about some encryption node.. But im sorry... I checked the workflow itself and I see nothing of the sort.
I just keep getting missing nodes even though it all loads without errors. Text Concatenate
Text Multiline fails
@salter883279 you can safely delete those two nodes, you don't need them at all. they only work for combining two text prompts as one. instead you can write your prompt inside the MinimaxH3 node's prompt section
@JustTrying2026 could you send me your workflow as a snapshot?
@reverentelusarca issue resolved by deleting the nodes. It at least works now. Thanks again for your help..
i know that some people are using for audio-only, i have not figured yet how.
Hey, I wrote a detailed article here, thanks for the inspiration: https://x.com/el_mejnun/status/2086212384503599165
Thanks for the flow 🤩
PS: tweaked it with turbo lora, sage att/spectrum apply and some upscale/postprocessing nodes - its included in this image: https://civitai.com/images/138951616

















