Image 1 — Fizgig 4.0 is out : Minimax H3 Combined Video File, Audio Files wav mp3 etc, Photo training in one dataset.  High Quality training samples (incl video) + turbo (finally) and new 'Gizmo' and AV dataset Prep tool. And Int 8 LARGE speedup for 16gb users.
Image 2 — Fizgig 4.0 is out : Minimax H3 Combined Video File, Audio Files wav mp3 etc, Photo training in one dataset.  High Quality training samples (incl video) + turbo (finally) and new 'Gizmo' and AV dataset Prep tool. And Int 8 LARGE speedup for 16gb users.
Image 3 — Fizgig 4.0 is out : Minimax H3 Combined Video File, Audio Files wav mp3 etc, Photo training in one dataset.  High Quality training samples (incl video) + turbo (finally) and new 'Gizmo' and AV dataset Prep tool. And Int 8 LARGE speedup for 16gb users.

Fizgig 4.0 is out : Minimax H3 Combined Video File, Audio Files wav mp3 etc, Photo training in one dataset. High Quality training samples (incl video) + turbo (finally) and new 'Gizmo' and AV dataset Prep tool. And Int 8 LARGE speedup for 16gb users.

I'll be making a Youtube video tomorrow for this. But I have one take away to share that I think is most important. H3, when you get the settings right is just fine with image based training without killing its video ability. Its even better when you combine photos and wavs, its super fast and you can train a voice with a dataset very easily. (I recommend shared trigger word). Video works too, but its slower, unavoidably. I'm not saying dont use it, its worth it for the right use cases. I'm just saying if you are not teaching the model anything new that photos and audio cant do, you are better with photos and audio. But when you do want to capture motion, it work very well. Anyway video coming tomorrow, with lots on Gizmo (the data set prep tool for video/audio) to make dataset prep easy. https://github.com/shootthesound/Fizgig

P.s the 16gb int8 speedup is from an an awesome community contribution from rintic-13 on Github.

u/shootthesound — 3 days ago

32:9 and 21:9 Wallpapers - via smart crop - 08/16 [OC]

This week its very surreal fantasy vibes. Hope you are all having a great weekend. All the high res and rest of the images (reddit compresses multi image posts) and crops on https://UltrawideWallpapers.net - Pete

u/shootthesound — 4 days ago
▲ 62 r/comfyui+1 crossposts

Fizgig - Rapid Minimax H3 LoRA training tutorial

This video includes all you need to train Minimax with both speed and high quality results.
Hit me up with comemtns, queries etc. Happy to do a style video also.
https://github.com/shootthesound/Fizgig

UPDATE: Pushed a vram optimisation for 16gb vram users that will speed up TE encoding at the start of training - Run the update bat to get it
UPDATE2: Additional fix out for 16gb users on pruned model - update to get it.

youtube.com
u/shootthesound — 8 days ago
▲ 47 r/MinimaxVideo+2 crossposts

ComfyUI-H3Studio for Single Node Long video Creation - Out Now

Crappy demo clip as I was short on time, but handy to see in context of the screenshot of this post

https://github.com/shootthesound/ComfyUI-H3Studio

Lots of hopefully clear instructions in the Github link and a basic example workflow.

Its my first time making a video editor after 15 years of using one every day, so there is a lot of carried over UX, and more I'll refine.

If you fancy it, this plays nicely with what is now a fast and high quality results Minimax Lora Trainer (getting good quality training to work in minimax has been a nightmare, but its there now): https://github.com/shootthesound/Fizgig

u/Hefty_Scallion_3086 — 8 days ago

New batch [7680x2160] Ultrawide Wallpapers (08/09) - link in comments for High res and crops

u/shootthesound — 11 days ago

32:9 and 21:9 Wallpapers - via smart crop - 08/09 [OC]

This week its all about post apocalyptic vibes and some Art Deco in the mix.. Happy Sunday btw! All the high res and rest of the images (reddit compresses multi image posts) and crops on https://UltrawideWallpapers.net - Pete

u/shootthesound — 11 days ago

New batch [7680x2160] Ultrawide Wallpapers (08/02) - link in comments for High res and crops

u/shootthesound — 18 days ago
▲ 15 r/comfyui+1 crossposts

Fizgig - Krea 2 Style LoRA LoKR Training Tutorial

Tutorial on how to train a Krea 2 Style Lora/LoKR with Fizgig on Windows (or Runpod / Linux)
https://github.com/shootthesound/Fizgig
This will work for 12gb vram and up. I've had many 1st hand reports in comments on Reddit/YT that its working for 8GB users too, but I've not personally tested that to claim it.

youtube.com
u/OneTrueTreasure — 19 days ago
▲ 44 r/comfyui+1 crossposts

Fizgig Rapid Krea 2 Lora Training Tutorial

Loads of comments, DMs, emails have asked for a video so here it is. The video for Krea 2 LoRA training includes Captioning, Adaptive Learning Rates, Per image and Per Epoch, Context Lora Mode, Automatic mid-train recaptioning and LR promotion/demotion for individual images and Likeness scoring in real-time.
https://github.com/shootthesound/Fizgig

youtube.com
u/shootthesound — 23 days ago

New batch [7680x2160] Ultrawide Wallpapers (07/26) - link in comments for High res and crops

u/shootthesound — 25 days ago

32:9 and 21:9 Wallpapers - via smart crop - 07/26 [OC]

Happy Sunday, latest drop is here as a lot of abstract fantasy and patterns. All the high res and rest of the images (reddit compresses multi image posts) and crops on https://UltrawideWallpapers.net - Pete

u/shootthesound — 25 days ago

Fizgig Krea 2 training features update

https://github.com/shootthesound/Fizgig

Intelligent trainer

- Per-image loss tracking with self-adapting training runs — every image gets its own verdict (easy / suspect / stuck / exhausted) and its own learning rate

- Auto-recaptioning: stuck images get their captions rewritten mid-run by Qwen3-VL from what's actually in the picture, then re-encoded and given a fresh start

- Auto-exclusion of unfixable images — after two failed recaption attempts a genuinely bad image is dropped from the run entirely, with safety rails so healthy images can never be excluded

- Problem Images window — live thumbnails, verdicts, and loss trends during training; edit a caption mid-run and it's picked up at the next epoch

- Adaptive learning rate that moves in both directions — probes up when loss is descending cleanly, backs off and rolls weights back when things go unstable

Dataset intelligence

- Look Consistency Filter — ArcFace face-embedding scoring of every dataset image against 3 baselines, catching identity drift that loss curves can't see

- Look-outlier warm-up — unusual-but-real images (profiles, tight angles) enter training gently at reduced LR and ramp up, instead of being punished or excluded

Live feedback

- Sample gallery with automatic likeness scoring — every preview scored against your dataset baselines on CPU while training runs, with a per-epoch trend chart and best-epoch highlight

- Training Run Visualiser — scrub your whole run epoch-by-epoch per prompt, export as WebM

Practical wins

- Train the full 12.9B RAW model on modest cards — fp8 residency + auto block-swap tuned to your GPU (~14 GB resident)

- Pause / Resume with zero quality loss — full optimizer, RNG, adaptive-LR and per-image-watch history restored, even across GUI restarts

- Context LoRA — train a new LoRA on top of an existing frozen one so they coexist at inference (no other trainer does this)

- Repair Studio — per-block sliders with live previews to fix an overbaked LoRA instead of retraining it

- ComfyUI-compatible output, no conversion step

u/shootthesound — 27 days ago
▲ 22 r/comfyui+1 crossposts

Image Save - Bling Edition for ComfyUI

https://github.com/shootthesound/ComfyUI-ImageSaveBlingEdition
The save node for ComfyUI reimagined. A session gallery on the node that survives restarts, hold-for-review triage that keeps your output folder clean, and one-click workflow recovery from any image ,plus formats, metadata, credits, auto mask side car images (mediapipe), watermarks, save to comfy inputs button etc, all remembered between sessions.

u/shootthesound — 28 days ago
▲ 83 r/comfyui+1 crossposts

KSampler Multi-Choice for ComfyUI

https://github.com/shootthesound/ComfyUI-KMS
See what your seeds have in mind before you spend the steps. Quick previews appear on the node, click your favourite and only that image gets rendered. You can click others after. Ideas welcome. Krea 2 workflow example in the node pack, but should work with any model. T2I and I2I supported. Cheers, Pete

u/shootthesound — 29 days ago