u/nickinnov

LTX 2.5 - Full-resolution workflows (no downscaling-upscaling)

LTX-2.5 is Lightricks' open video generation model and once again they have taken the Comfy UI image-to-video workflow and applied their downscaling-rendering-upscaling technique presumably so it runs faster and works on lower-spec hardware - which is fair enough.

But for those of us who invested in Jensen Huang's next leather jacket by buying a DGX Spark can use the extra memory room for rendering at full resolution.

Here are the workflows, adapted from the original Comfy UI LTX 2.5 templates and working with the same models:

LTX 2.5 Image to Video Full Resolution

https://cdn.lansley.com/comfyui-assets/LTX%202.5%20Image%20to%20Video%20FullRes.json

LTX 2.5 Text to Video Full Resolution

https://cdn.lansley.com/comfyui-assets/LTX-2.5%20Text%20to%20Video%20FullRes.json

These workflows are set to the distilled BF16 version of LTX 2.5 but work just as well with the distilled 'int8-convrot' version in terms of better detail in the full resolution versions compared to the templates supplied by Comfy.

The difference is most noticeable in Image to Video when the you want the action to depart significantly from the supplied first frame. The supplied template struggles to re-apply the detail during the upscaling section of the workflow whereas the full-res version keeps the detail in every frame especially when new content has to be invented that was not in the first frame.

The Text to Video workflow includes a prompt about two guys playing ball on a beach involving the need for detailed sea surf and sand resolution, which was noticeably better in this full resolution version compared to the supplied template workflow using the same prompt ('enhance prompt' tuned off).

The performance of the full resolution versions are of course much slower.

Supplied reduced-res template: 8 x 8 seconds + 3 x 33 seconds = 163 seconds.

That's 8 x low-res rendering steps then 3 x upscale steps.

Full-res text to image is 8 x 33 seconds = 264 seconds (rendering steps only of course).

As ever, trust nothing you ever download and make sure the flowchart JSON files look OK before running them in ComfyUI. Otherwise, enjoy! Would be great to get your feedback.

reddit.com
u/nickinnov — 2 days ago
▲ 17 r/LTXvideo+1 crossposts

LTX 2.5 - Full-resolution workflows (no downscaling-upscaling)

LTX-2.5 is Lightricks' open video generation model and once again they have taken the Comfy UI image-to-video workflow and applied their downscaling-rendering-upscaling technique presumably so it runs faster and works on lower-spec hardware - which is fair enough.

But for those of us who invested in Jensen Huang's next leather jacket by buying a DGX Spark can use the extra memory room for rendering at full resolution.

Here are the workflows, adapted from the original Comfy UI LTX 2.5 templates and working with the same models:

LTX 2.5 Image to Video Full Resolution

https://cdn.lansley.com/comfyui-assets/LTX%202.5%20Image%20to%20Video%20FullRes.json

LTX 2.5 Text to Video Full Resolution

https://cdn.lansley.com/comfyui-assets/LTX-2.5%20Text%20to%20Video%20FullRes.json

These workflows are set to the distilled BF16 version of LTX 2.5 but work just as well with the distilled 'int8-convrot' version in terms of better detail in the full resolution versions compared to the templates supplied by Comfy.

The difference is most noticeable in Image to Video when the you want the action to depart significantly from the supplied first frame. The supplied template struggles to re-apply the detail during the upscaling section of the workflow whereas the full-res version keeps the detail in every frame especially when new content has to be invented that was not in the first frame.

The Text to Video workflow includes a prompt about two guys playing ball on a beach involving the need for detailed sea surf and sand resolution, which was noticeably better in this full resolution version compared to the supplied template workflow using the same prompt ('enhance prompt' tuned off).

The performance of the full resolution versions are of course much slower.

Supplied reduced-res template: 8 x 8 seconds + 3 x 33 seconds = 163 seconds.

That's 8 x low-res rendering steps then 3 x upscale steps.

Full-res text to image is 8 x 33 seconds = 264 seconds (rendering steps only of course).

As ever, trust nothing you ever download and make sure the flowchart JSON files look OK before running them in ComfyUI. Otherwise, enjoy! Would be great to get your feedback.

reddit.com
u/Hefty_Scallion_3086 — 4 days ago