Turn portraits into sharp, cinematic close-up videos with native audio.
Who it's for: creators who want this pipeline in ComfyUI without assembling nodes from scratch. Not for: one-click results with zero tuning - you still choose inputs, prompts, and settings.
Open preloaded workflow on RunComfy
Open preloaded workflow on RunComfy (browser)
Why RunComfy first
- Fewer missing-node surprises - run the graph in a managed environment before you mirror it locally.
- Quick GPU tryout - useful if your local VRAM or install time is the bottleneck.
- Matches the published JSON - the zip follows the same runnable workflow you can open on RunComfy.
When downloading for local ComfyUI makes sense - you want full control over models on disk, batch scripting, or offline runs.
How to use (local ComfyUI)
1. Load inputs (images/video/audio) in the marked loader nodes.
2. Set prompts, resolution, and seeds; start with a short test run.
3. Export from the Save / Write nodes shown in the graph.
Expectations - First run may pull large weights; cloud runs may require a free RunComfy account.
Overview
Turn your portrait into a cinematic, face-forward video. Get natural facial motion and synchronized native audio. Frame fashion, dialogue, or social clips tightly. Two latent stages preserve detail. 2x upscaling sharpens each frame. Start fast with this ready-to-run LTX setup.
Important nodes:
Key nodes in Comfyui LTX 2.3 Close-Up Shots workflow
LTXVImgToVideoInplace(#161 and #160)
Anchors the portrait into the latent video so the face remains stable while subtle push-ins and micro-movements play out. Keep it active for image-to-video; switch itsbypassonly when using text-to-video mode. For the best LTX 2.3 Close-Up Shots results, start from a sharp, evenly lit reference and avoid occlusion over the mouth.LTXVLatentUpsampler(#118)
Applies the native LTX 2.3 x2 spatial upscaler in latent space before the second-pass refinement. This preserves motion coherence while increasing detail. For close-ups, keeping the upscale at x2 typically offers the best balance of sharpness and stability.CFGGuider(#129 and #103)
Combines your positive and negative conditioning into a unified guidance signal for both modalities. Small adjustments to guidance strength can tighten or loosen adherence to your script and framing. If you increase guidance, consider slightly softening the negative list to avoid over-constraining expressions.SamplerCustomAdvanced(#113 and #119)
The first pass favors a speed-oriented strategy to establish motion and audio quickly, while the second pass switches to a fidelity-oriented strategy to sharpen details. If you change sampler strategy, keep the first pass fast and the second pass precise so LTX 2.3 Close-Up Shots maintains its cinematic polish.LTX2_NAG(#342)
Introduces noise-aware guidance that balances video and audio conditioning. Raising the video emphasis can firm up framing and pose; raising the audio emphasis can harden lip sync. Tune conservatively and in tandem withCFGGuiderto avoid over-constraint.LTXVConditioning(#107 and #22)
Binds your prompts and frame rate to the latent timeline. Maintain a single, scene-level positive prompt and a compact negative prompt for the cleanest alignment. The attached frame rate ensures speech pacing and motion cadence stay in lockstep.
…
Notes
LTX 2.3 Close-Up Shots in ComfyUI - Audio and 2x Upscale - see RunComfy page for the latest node requirements.
Description
Initial release - LTX-2.3-Close-Up.
