CIVITAI / Workflows

LTX 2.3 Video Extension OmniNFT + Relay Vertical Widening Workflow

Watch the full video first if you want to understand how this LTX 2.3 video extension workflow works in practice. The video shows how a vertical video can be expanded into a wider frame, why OmniNFT + Relay guidance matters, and how to launch the workflow online without rebuilding the full ComfyUI setup locally. This ComfyUI workflow is designed for LTX 2.3 video extension, vertical-to-wide frame expansion, and guided video outpainting. The main purpose of this workflow is to take a narrow or vertical video-style input and expand it into a wider composition while keeping the original subject, motion, and visual identity as stable as possible. Instead of simply stretching the image or cropping the video, this workflow uses LTX 2.3 generation to synthesize new side areas and create a more natural wide-frame result. The workflow is built around the LTX 2.3 distilled 1.1 route. It uses ltx-2.3-22b-dev-dare-merged-distilled-1.1.safetensors as the main checkpoint, Gemma-based text encoding, LTX Audio VAE, LTXVConditioning, LTXAddVideoICLoRAGuide, LTXVCropGuides, LTXVConcatAVLatent, LTXVSeparateAVLatent, ManualSigmas, CFGGuider, SamplerCustomAdvanced, VAEDecodeTiled, and VHS_VideoCombine. The graph also includes image resizing, color correction, image switching, and image concatenation nodes, which are important for comparing and assembling the expanded output. The key idea is guided extension. The source frame or reference image is first prepared through resizing and LTXVPreprocess. Then LTXAddVideoICLoRAGuide injects the visual guide into the LTX generation process, helping the model preserve the original content while expanding beyond the initial frame boundary. LTXVCropGuides helps manage the guided area so the model can focus on the extension region instead of freely changing the whole image. The workflow also uses audio-video latent logic. Empty video latent and empty audio latent are created, then combined through LTXVConcatAVLatent before sampling. After generation, LTXVSeparateAVLatent separates the video and audio latent streams again. This makes the workflow compatible with LTX 2.3 audio-video generation structure and final video output. Compared with ordinary video resizing, this workflow does not only change the canvas size. It generates new visual content for the expanded region. Compared with basic image outpainting, it works in a video pipeline

LTXV 2.3 #character
View original on Civitai
LTX 2.3 Video Extension OmniNFT + Relay Vertical Widening Workflow

Public versions

v1.0

LTXV 2.3