This workflow extends the MiniMax H3 digital human setup to a two-character dialogue scene. It is built for controlled short drama, paired avatars, character interaction, and audio-driven conversation where each speaker needs distinct timing, expression, and visual role separation.
The workflow uses the MiniMax H3 local ComfyUI graph with the connected model, VAE, encoder, sampler, resolution, and final video output chain shown in the JSON.
Main features:
- MiniMax H3 dual-character reference-to-video route
- Connected MiniMax H3 Ref2VA diffusion model with Qwen3-VL text encoding
- Separate character references can define Subject 1 and Subject 2
- Scene reference can be used to keep the environment and visual tone consistent
- Audio-driven timing supports staged speaker turns and dialogue beats
- 16:9 widescreen output at the connected 0.9 megapixel resolution route
- Final video combine route exports the generated scene as an audio-video result
Suggested workflow:
Assign each input a strict role before generation: one reference for the first character, one for the second character, and an optional scene reference for the environment. In the prompt, define who speaks during each time range and who stays silent. This helps reduce speaker confusion, mouth movement drift, and unwanted duplicate characters.
RunningHub Workflow
Try the workflow online right now - no installation required.
Workflow: https://www.runninghub.ai/post/2085408407900762114?inviteCode=rh-v1111
If the results meet your expectations, you can later deploy it locally for customization.
Fan Benefits: Register to get 1000 points + daily login 100 points - enjoy 4090 performance and 48 GB super power!
Bilibili Updates (Mainland China & Asia-Pacific)
If you're in the Asia-Pacific region, you can watch the video below to see the workflow demonstration and creative breakdown.
Bilibili Video: https://www.bilibili.com/video/BV1kUuW6FEVs/
Support Me on Ko-fi
If you find my content helpful and want to support future creations, you can buy me a coffee.
Every bit of support helps me keep creating.
Ko-fi: https://ko-fi.com/aiksk
Business Contact
For collaboration or inquiries, please contact aiksk95 on WeChat.
打开下方链接即可在线体验,无需安装。
工作流:https://www.runninghub.ai/post/2085408407900762114?inviteCode=rh-v1111
如果你觉得效果理想,也可以在本地进行自定义部署。
粉丝福利:注册即可领取 1000 积分,每日登录再领 100 积分,体验 4090 和 48 GB 大显存性能。
B站视频(中国大陆及亚太地区)
如果你在中国大陆或亚太地区,可以通过下面的视频查看工作流的实测效果与创作思路。
B站视频:https://www.bilibili.com/video/BV1kUuW6FEVs/
我会在夸克网盘持续更新模型资源:
https://pan.quark.cn/s/07bdc81784ce
Description
This workflow extends the MiniMax H3 digital human setup to a two-character dialogue scene. It is built for controlled short drama, paired avatars, character interaction, and audio-driven conversation where each speaker needs distinct timing, expression, and visual role separation.
The workflow uses the MiniMax H3 local ComfyUI graph with the connected model, VAE, encoder, sampler, resolution, and final video output chain shown in the JSON.
Main features:
- MiniMax H3 dual-character reference-to-video route
- Connected MiniMax H3 Ref2VA diffusion model with Qwen3-VL text encoding
- Separate character references can define Subject 1 and Subject 2
- Scene reference can be used to keep the environment and visual tone consistent
- Audio-driven timing supports staged speaker turns and dialogue beats
- 16:9 widescreen output at the connected 0.9 megapixel resolution route
- Final video combine route exports the generated scene as an audio-video result
Suggested workflow:
Assign each input a strict role before generation: one reference for the first character, one for the second character, and an optional scene reference for the environment. In the prompt, define who speaks during each time range and who stays silent. This helps reduce speaker confusion, mouth movement drift, and unwanted duplicate characters.
RunningHub Workflow
Try the workflow online right now - no installation required.
Workflow: https://www.runninghub.ai/post/2085408407900762114?inviteCode=rh-v1111
If the results meet your expectations, you can later deploy it locally for customization.
Fan Benefits: Register to get 1000 points + daily login 100 points - enjoy 4090 performance and 48 GB super power!
Bilibili Updates (Mainland China & Asia-Pacific)
If you're in the Asia-Pacific region, you can watch the video below to see the workflow demonstration and creative breakdown.
Bilibili Video: https://www.bilibili.com/video/BV1kUuW6FEVs/
Support Me on Ko-fi
If you find my content helpful and want to support future creations, you can buy me a coffee.
Every bit of support helps me keep creating.
Ko-fi: https://ko-fi.com/aiksk
Business Contact
For collaboration or inquiries, please contact aiksk95 on WeChat.
打开下方链接即可在线体验,无需安装。
工作流:https://www.runninghub.ai/post/2085408407900762114?inviteCode=rh-v1111
如果你觉得效果理想,也可以在本地进行自定义部署。
粉丝福利:注册即可领取 1000 积分,每日登录再领 100 积分,体验 4090 和 48 GB 大显存性能。
B站视频(中国大陆及亚太地区)
如果你在中国大陆或亚太地区,可以通过下面的视频查看工作流的实测效果与创作思路。
B站视频:https://www.bilibili.com/video/BV1kUuW6FEVs/
我会在夸克网盘持续更新模型资源:
https://pan.quark.cn/s/07bdc81784ce
minimax h3
workflows
ai video
runninghub
workflow
comfyui
lip sync
talking avatar
audio driven video
two character dialogue
dual digital human
Details
Downloads
41
Platform
CivitAI
Platform Status
Available
Created
8/7/2026
Updated
8/8/2026
Deleted
-
