CIVITAI / Workflows

LTX 2.3 Dual Digital Human | IC Edit No-Subtitle Dialogue Workflow

This workflow is designed for LTX 2.3 dual-person digital human dialogue generation, with IC Edit-style control and a strong focus on clean subtitle-free output. Its main purpose is to take a two-person reference image or character scene, generate a controlled dialogue-style video, and keep the final result clean without unwanted subtitles, fake captions, random text, watermark-like marks, overlays, or UI-style artifacts appearing on the screen. Compared with a single-person digital human workflow, this setup is more demanding because it needs to maintain two character identities at the same time. A good dual-person dialogue video must preserve left-right placement, facial consistency, clothing, body proportion, camera framing, background stability, and interaction logic. If the workflow is not controlled well, the two characters may swap positions, merge faces, duplicate body parts, drift away from the original image, or create random mouth movement that does not match the intended dialogue structure. The workflow uses LTX 2.3 as the main video generation backbone, with LTX video VAE, LTX audio VAE, image resizing, image-to-video conditioning, LTXVConditioning, LTXVImgToVideoConditionOnly, LTXVConcatAVLatent, LTXVSeparateAVLatent, SamplerCustomAdvanced, ManualSigmas, latent upscaling, tiled VAE decoding, audio decoding, CreateVideo, SaveVideo, and VRAM cleanup logic. This makes it a more complete production workflow rather than a simple one-pass image animation graph. The core generation design follows a staged rendering structure. The first stage builds the base motion, character presence, camera structure, and dialogue performance from the reference image and prompt conditioning. Later stages continue from the generated latent result with lower sigma values, refining motion stability, facial detail, clothing texture, background consistency, and final visual quality. This staged approach is especially useful for two-person digital human scenes because both subjects need to remain coherent across the full video. A key feature of this workflow is its dual-character control direction. The workflow is built for restrained dialogue performance rather than chaotic motion. The ideal output should show two people facing the camera or interacting naturally, with subtle head movement, mouth movement, facial expression changes, small hand gestures, and stable bod

LTXV 2.3 #character
在 Civitai 查看原始条目
LTX 2.3 Dual Digital Human | IC Edit No-Subtitle Dialogue Workflow

公开版本

v1.0

LTXV 2.3