H3 + LTX 2.5 - Upscaler... before H3-Regenerate-2k arrives

WORKFLOW: https://pastebin.com/raw/xPTwRh9P (H3 not embedded works for any video - mainly for H3 upres with LTX 2.5 INT8 Distilled / SageAttn / SolAttention)

This was my very first gen from H3, time to revisit and upscale it with LTX 2.5 + Spatial 2x. If you notice - when she's far away "walk" it's better! This is really great. Of course, Spatial 2x applies some softness here and there, overall is not bad. Didn't test - SeedVR2 and FlashVSR yet

Perhaps, we can still use LTX 2.5 for now for upscaling (including: voice dub, LTX director, connecting/extending shots), and make gems gens with H3.

Till new beast arrives.... Minimax H3-Regenerate-2K (probably will be superior to anything).

H3 original (super low res): 864 x 480 / 18 Steps / Pruned INT8
Spatial 2x LTX 2.5: 1728 x 960 (Distilled INT8, render time pm: 338 sec - at RTX 4090 / 128gb RAM / SageAttn + SolAttn)

Ps. On phone maybe hard to see - checkout on desktop
Cheers

u/No_Damage_8420 — 9 days ago
▲ 90 r/LovingAIVisuals+1 crossposts

LTX 2.5 vs H3 - Rundown Multi-Angle Scene

HI everyone,
Today we got release of LTX 2.5, so lets compare with more advance scenes along side Minimax H3.

3 videos in order - same PROMPT (with <Shot 1>...<Shot 2>)

  1. LTX 2.5 1-stage (8 step distilled INT8)
  2. LTX 2.5 2-stage (8 step + 3 step upscaler default ComfyUI template distilled Int8)
  3. Minimax H3

We can definitely say, LTX is upgraded (more detailed textures vs LTX 2.3...lots of snow, snow slush etc), is it better then H3? Judge yourself ;)
cheers

Here we go PROMPT:
[Shot 1] Handheld selfie-style video shot on a smartphone at 60fps, capturing a bleak, emotional scene under the harsh fluorescent lights of a drafty, concrete train station concourse. A young, strikingly beautiful girl—around eighteen years old—sits slumped against a cold tiled wall in a corner. She holds the smartphone in a trembling right hand, filming herself and her surroundings. She wears a thin, threadbare denim jacket covered in dark grease stains and road salt, a stretched-out oversized sweater torn at the collar, and frayed canvas pants soaked with melting slush at the cuffs. Clustered beside her in the freezing draft is an old German Shepherd with a matted, dirt-streaked coat, shivering against her knee. Heavy snow and howling wind blur through the open station doors behind them. Tears stream down her smudge-streaked cheeks as she speaks into the phone in a quiet, trembling voice: <d>[English with a soft, heartbroken voice] It's so cold tonight... please, nobody even looks at us... we just need a little food...</d>. At 00:05:500 Cut to [Shot 2] Low-angle handheld shot from her perspective on the floor. Commuters in thick, clean winter coats and boots hurry past her in a fast blur, heads turned away, completely ignoring her outstretched, gloveless left hand. Her breath forms dense white clouds in the freezing air. The German Shepherd lets out a soft, whimpering whine and rests its heavy head on her lap. She wipes a tear from her nose with her sleeve, her voice cracking with desperation: <d>[English with choked, sobbing breaths] Please... just a piece of bread for him... anything...</d>. At 00:11:000 Cut to [Shot 3] Upward camera angle as a middle-aged man in a dark wool overcoat and scarf suddenly stops in front of her. His boots come to a halt in the wet slush beside her dog. He crouches down to her eye level, looking at her and the shivering German Shepherd with genuine concern and warmth. He gently reaches into his coat pocket, pulling out a warm paper bag from a bakery and a thermal travel mug, speaking in a gentle, compassionate voice: <d>[English with a warm, caring tone] Hey... hey, don't cry. Here, take this—it's hot soup and fresh bread. Are you okay?</d>. At 00:16:000 Cut to [Shot 4] Close emotional selfie framing as her eyes widen in tearful disbelief. She hugs the warm paper bag to her chest with both hands, tears pouring down her face as she smiles through her sobs, looking up at the man and then into the camera lens: <d>[English with a tearful, weeping whisper] Thank you... oh god, thank you so much... bless you...</d>. The German Shepherd gently licks her cold hand as the man reaches out to pet the dog's head, cutting the video to black on a powerful, dramatic note at 00:20:000, overall_soundscape: Howling blizzard winds outside open station doors, heavy footsteps echoing on wet tile floors, distant train arrival announcements, shivering whimpers from the dog, and her quiet, heartbreaking sobs, non_diegetic_music: Soft, somber cinematic violin pads fading in subtly under the diegetic audio to heighten the emotional drama

u/Koala_Confused — 9 days ago
▲ 242 r/NeuralCinema+1 crossposts

Minimax H3 ~ "Hack" 50+ reference or more

I discover this by playing around, so good news is - we can have much more references then 9.

  1. "ref2v" - adjust "ref_image_size" to MAX
  2. Use just 1 IMAGE REFERENCE with - many cutouts, items, elements mentioned at once in PROMPT:
  3. https://preview.redd.it/minimax-h3-hack-50-reference-or-more-v0-revg36ju4fih1.png?width=1957&format=png&auto=webp&s=027debaf03d7c630f925170578e6b83c552a7096
  4. Write simple prompt, reference image once, all other elements H3 will combine into final scene:

&lt;Picture 1&gt; woman in &lt;Picture 2&gt; luxury bathroom is touching her face showing her silver earrings, camera slow motion up-close on face and torso, she puts on glasses, looks at mobile purple phone puts to her ear and smiles to camera.

Once again H3 it's beyond amazing....
Also by including multiple faces - different expressions H3 learns expressions etc. teeth, looks, in a away - we don't need character LORA.
I guess this is great find for all of us.
Cheers

u/No_Damage_8420 — 11 days ago

Minimax H3 - 40% speed-up / Quality Holds (no turbo LORA)

*UPDATE*
Even faster with - https://www.reddit.com/r/StableDiffusion/comments/1vhlfmw/minimax_h3_firstblockcache_for_comfyui_3033_lower/
You can even combine them both for extreme speeds

Hey all,
Lots of addons been released by hours, one interesting speed-up is - Spectrum (cannot be used with EasyCache): https://github.com/xmarre/ComfyUI-Spectrum-MiniMax-H3

Follow more here:
https://www.reddit.com/r/StableDiffusion/comments/1vf1ze3/spectrum_acceleration_for_minimax_h3_in_comfyui/?utm_source=chatgpt.com

Video above rendered with Spectrum / 20 Steps
cu130+SageAttn + rtx 4090, 128gb RAM, Win11

Spectrum: 201 seconds
Default: 320 seconds

Quality still holds, I was testing all current H3 turbo LORA's - all seems to kill details and/or severely degrade sound. I'm sticking with 20 steps - solid.
Cheers

u/No_Damage_8420 — 14 days ago

Minimax H3 - extreme, uncensored, imagination is your limit

H3 brings lots of creative power, some more extreme showcase. Grab weights while you can.
Cheers

u/No_Damage_8420 — 16 days ago

LTX 2.3 Full 360 angle CAMERA Control - via CrossView-Warp IC-LoRA

Hi all,

In the end Full 360 angle camera control is possible via LTX 2.3 CrossView-Warp.

After further experimenting with amazing new LTX Lora - CrossView Warp ( https://www.reddit.com/r/StableDiffusion/comments/1v89kih/ltx_crossviewwarp_iclora_change_the_camera_angle/ ) , seems Lora is able to do full 360 renders :) 3D Model (via Pixal3D in my case, probably Hy-World or Apple Sharp would work too etc.) used was STATIC, no animation of any kind, however, seems strong "anchor point" for LORA to extract/read rotations/motions.

Process here was more complicated just to see if works, proof of concept.
Hitman image -> Scail-2 -> Pixal3D (full model extraction for manual 360 rotation) -> LTX 2.3 + CrossView Warp Lora and we have great result.

Big thanks goes to creator: DryDream6994 of this amazing Lora, we hope for new releases.
Video looks cartoonish, well we using HITMAN game character in photoreal downtown setting.
Onto more research, Cheers

https://preview.redd.it/350d42u68bgh1.png?width=2849&format=png&auto=webp&s=0c9197ef319f79be416c694d9b063711fffb8d3d

https://preview.redd.it/sey0b8ig8bgh1.png?width=3324&format=png&auto=webp&s=5007ace62c1edcff40230b6543cd4ba3b0821947

u/No_Damage_8420 — 22 days ago
▲ 220 r/NeuralCinema+1 crossposts

LTX CrossView-Warp IC-LoRA - Change the camera angle and orbiting path of an existing video more precisely

Hello Everyone! Let me share my newest camera control IC-LoRA where you can define the new camera angle on an orbit sphere instead of using just text prompt.

You can download the model from here: https://huggingface.co/Cseti/LTX2.3-22B_IC-LoRA-CrossView-Warp
You'll need this custom node to be able to define the new camera angle or path: https://github.com/cseti007/ComfyUI-CrossViewWarp

You can find example workflow in the custom node's "example_workflow" folder.

Have fun!

u/DryDream6994 — 24 days ago
▲ 604 r/NeuralCinema+1 crossposts

LTX Director - An All-In-One Timeline Editor. I2V, T2V, FLFF, Prompt Relay, Custom Audio, and more! Unlock LTX 2.3's full potential!

LTX Director is a timeline editor that allows you to easily compose LTX videos. It is the evolution of my previous nodes, LTX Sequencer and Multi Image Loader, and will hopefully help unlock the huge potential of LTX 2.3.

Download for free here: https://github.com/WhatDreamsCost/WhatDreamsCost-ComfyUI

I worked on this for 6 days straight, spending 16+ hours a day vibe coding it with Gemini. Hopefully it helps you create cool stuff easier!

Main Features:

  • Fully Functional Timeline Editor: Add image, text, and audio segments to control exactly what happens and when. Easily trim, cut, and edit segments with a (hopefully) intuitive interface.
  • Prompt Relay integrated: This unlocks the ability to have granular control over video generation. For more information on Prompt Relay go here, https://gordonchen19.github.io/Prompt-Relay/
  • First, Middle, Last Frame Support: This node has by far the easiest method of creating first/last frames videos. It supports any number of keyframes, and will be the successor of my previous nodes.
  • Custom Audio Support: Import, trim, and combine your own audio clips in this node. Enabling custom audio is as simple as clicking 1 button. It is also compatible with every other feature in the node, include first/last frames, t2v, i2v, and prompt relay.
  • Image to Video: Part of the goal of this node was to make it easier to do everything, including Image to Video. It has built in resize functionality, and of course all the benefits of the prompt relay and custom audio integration.
  • Text to Video: Simply load any images and use text segments to create T2V videos. Compatible with all other features of the node.
  • And more much! I'm only scratching the surface, but this really does allow you to create shots that were almost impossible (if not impossible) to do normally with LTX 2.3.
u/No_Damage_8420 — 3 months ago