r/texttovideo

HVAC / MOLD Company Ai Text To Video (Help me determine best software)

Good Day!

My boss is requesting that our team generates Ai Videos for our HVAC / Mold Company. The average time per video is about 45 seconds to 1 minute. I want to find a good quality software yet cost effective.

May I ask for recommendations? I have been checking so many forums, threads, posts, etc. and each person has positive and negative comments and it's hard to just try so many softwares since I would need to spend money to actually test it. For those that have any suggestions, please let me know. Thank you

reddit.com
u/Damien_Aurelius — 1 day ago

What's the best AI image generator for e-commerce visuals?

Anyone here tried AI tools for making product images? Need something clean and professional for an online store. Heard Magnific has some AI stuff, but not sure if it's top tier for e-comm visuals. Thoughts?

reddit.com
u/isomorphix19 — 6 days ago

Any good free AI tools for real-time image generation?

Lately, i've been diving into AI generated art and want to try some tools that let me create images on the fly without paying. There are so many options out there, but I keep running into either super limited free versions or ones that require cloud credits. I stumbled on something called Magnific, which seems promising for quick image generation, but haven’t tested it much yet. Anyone familiar with free AI services that can do real-time image generation without crazy restrictions? Would love to avoid the heavy watermarking some tools slap on tha free outputs.

reddit.com
u/amyyrosse_ — 7 days ago

How to create a professional storyboard using AI?

So I recently got tasked with puttting together a storyboard for a client video, and honestly, I'm drowning in all the options. I figured AI tools might speed things up, but I have zero experience with any storyboard specific software. Is there a good way to actually create a professional looking storyboard using AI, or am I just kidding myself?

reddit.com
u/reddit_lurker1234567 — 8 days ago

Need help finding the best alternative to Midjourney

So, I want to try something new but keep hitting walls with Midjourney’s limits. Anyone know a good alternative that’s user-friendly and looks proffesional? Came across Magnific and not sure if they're on the same level.

reddit.com
u/LadyDemura — 9 days ago

looking for an ai video tool that supports actual partial edits

i finally get a clip that mostly works and then i notice one prop is wrong or a character's outfit shifted color halfway through and the only fix most tools give me is regenerating the whole thing from scratch. sometimes it comes out worse the second time, sometimes it fixes the one thing and breaks three others. feels like im gambling every time i hit generate instead of editing. has anyone dealt with this kind of problem and found something that handles partial edits properly?

a little update. heard dreamina has something called seedance 2.5 that supposedly lets you edit specific segments without touching the rest of the clip, up to five timestamped edits per clip from what ive read. havent tried it myself yet but planning to give it a shot since it sounds like exactly what id need. 

reddit.com
u/Awa-Hilgendorf — 11 days ago
▲ 24 r/texttovideo+3 crossposts

Comfy Org invited Minimax H3 team to talk: text summary of the stream

(Summary made with help of AI as you must expect).

Original video: https://www.youtube.com/watch?v=S9O3FPumX4Q

MiniMax H3 (Hailuo 3) — Video Generation Model

The video introduces the open-weight release of the MiniMax H3 (Hailuo 3) video generation model. H3 is a 60-billion-parameter model capable of text-to-video, image-to-video, reference-to-video, in-place editing, and native audio generation.

Because it is open-weight, the community has already integrated it deeply into ComfyUI, allowing for complex, multi-modal video generation on local machines.

Prompting Techniques & Best Practices

The hosts and creators shared several key strategies for getting the best results out of H3:

  • Keep It Straightforward: The model understands direct, straightforward language very well. You don't necessarily need overly complex "prompt engineering" jargon.
  • Use an LLM as a "Prompt Enhancer": In the ComfyUI workflow, Comfy Rob uses an LLM node to take a basic concept and expand it into detailed, shot-by-shot prompts. This helps inject specific shots into the generated video seamlessly.
  • Dialogue in Quotes for Lip-Sync: If you are using reference audio and want a character to speak, put the exact dialogue in quotation marks ("") within your text prompt. The model will automatically sync the character's lip movements to the referenced audio file.
  • Context IR API (Intermediate Representation): If you are working with the API or advanced nodes, H3 has a feature that optimizes your prompt based on the multiple reference images you provide, helping the model figure out how to stitch different modalities together.

ComfyUI Workflow Techniques

Comfy Rob demonstrated a Reference-to-Video workflow, which you can find in the ComfyUI Template Library by searching for "Minimax." Here are the technical tips for setting up your nodes:

  • Resolution / Megapixels Setting: Instead of traditional width/height settings, you often set the "Megapixels" node. Rob recommends starting at 0.4 megapixels, which outputs a resolution of roughly 864 × 480p.
  • Controlling Shots and Timing: You can define the number of shots and the total length of the video. For example, if you want a 10-second video and set it to 5 shots, the model will generate a new shot every 2 seconds.
  • Multi-Image Referencing: H3 allows up to 12 reference inputs. You can input multiple angles of a product (such as the earbuds example shown) to maintain strong temporal and spatial consistency across different generated shots.

Examples of What the Model Excels At

If you are looking for inspiration for your prompts, the video showcased three main areas where H3 currently excels:

1. Cinematic Product Advertisements

Rob used images of earbuds to create a sleek, professional 10-second commercial with changing camera angles, demonstrating the model's high consistency with product references.

2. Music Videos & Lip-Syncing

A video of a multi-eyed alien at a post office was shown where the alien's movements perfectly matched the rhythm of a song, and it sang the lyrics with highly accurate lip-syncing.

3. Complex Physics — Cloth & Fur

A video of a cat moving under a thick blanket demonstrated the model's impressive grasp of cloth physics, weight, and fur consistency without morphing or noticeable artifacts.

Hardware & Optimization Tips

Crucial for Local ComfyUI Users

H3 is a massive model, requiring approximately 120 GB of VRAM natively, so running it on consumer GPUs requires some optimization techniques:

  • Use Kijai's LoRA: The community member Kijai released a 4-to-8-step LoRA. Adding this to your workflow can drastically speed up generation times on lower-end hardware.
  • Use Quantized Models: Make sure you download quantized versions of the model, such as GGUF or FP8 versions, built for ComfyUI.
  • Update ComfyUI: Ensure you are running the absolute latest version of ComfyUI. It uses fine-grained offloading specifically for H3, automatically moving parts of the calculation between your system RAM and your GPU's VRAM. This makes it possible to run the massive model on a standard 24 GB GPU, such as an RTX 3090 or RTX 4090.
u/Hefty_Scallion_3086 — 12 days ago