Image 1 — Trying out LoKr instead of LoRA on Krea2
Image 2 — Trying out LoKr instead of LoRA on Krea2
Image 3 — Trying out LoKr instead of LoRA on Krea2

Trying out LoKr instead of LoRA on Krea2

Dataset of 43 images, captioned with qwen3 VL 4B instruct,

50 word caption focusing on: Composition, Subject's hair, expression, clothes, pose, background

Training Parameters: (10 rep x 43 image) x 6 epoch, saved every 430 steps.

resolution: 512, 768, learning rate 0.0001, Optimizer: Automagic v2,

cache text embeddings, cache latents,

My setup is RTX 4070 super + 64gb RAM + Paging file size: 65536

I had to offload my text encoder 100% and Transformer 75%

Each LoKr tested with same caption + Noise on Krea2Turbo, i trained some character loras before, this feels very similar to the LoRAs. Not many ground breaking improvements, but got decent results from early 2k steps.

u/Beautiful_Egg6188 — 24 days ago

Made another Style LoRA on Krea2

https://civitai.red/models/2787790/artstyle-kuvshinovilya
Used the base/Raw version of the model on AIToolkit, 2220 steps total, resolution was set to 512, 768. Dataset had 37 images total.

all the images had simple caption about the subject and background, each caption under 50 words. No style described.
So works pretty well even without the trigger word "Kuvshinov_Ilya"
If using with other 3d styles, or realistic characters, its best to use this in your prompt. "Kuvshinov_Ilya, clean sharp line art, Soft focus, low bloom effect"

Took me 3.5 hours to train the model on rtx 4070 super + 64GB Vram. krea 2 Raw Model offloading set to 73% (works for me) on float8. everything else was set to default. Cached captions.

The current version is stronger one, it might push the characters to more feminine look, but a softer 1100 step version will be added soon. i have only tested a few prompts and it works much better for male characters.

u/Beautiful_Egg6188 — 1 month ago

2nd attempt at LTX2.3 + Krea2

Shots generated with Krea2+ my lora. Tried moving camera with more motion this time. LTX2.3 can barely adhere to the prompt. does random things, bad motion, and hand movement. took multiple attempts to only get one decent shot to use. Some of them I'm still not very pleased with.
Used Default ComfyUI workflow, used the "--novram" inside bat file.
my setup is rtx4070Super 12GB+ 64GB ram, Virtual Memory set to 64GB as well.
https://www.reddit.com/r/StableDiffusion/comments/1uxfwrw/havent_used_a_model_this_much_since_flux1dev/

u/Beautiful_Egg6188 — 1 month ago

IDG4 is perfect for moodboard.

I can reuse hex codes, Aesthetic, and photo description to get a similar vibe over multiple images. An open source model provides much better control over all the other closed models.

u/Beautiful_Egg6188 — 2 months ago

CRT screen game on Ideogram4 on CivitAI

each image took about 2 minutes, 20 steps, 1920x1088, on rtx 4070super+ryzen7700+64gb ram
all prompts are json formatted.

u/Beautiful_Egg6188 — 2 months ago

tested IdG4 against ZiT's strong suit (realistic portrait)

left Ideogram 4 fp8, right Zit. the generation still takes too long for IdG (70sec for turbo, 2mp. Zit does it in 30 sec)
ZiT just feels more natural looking compared to Ideogram4. But the control over characters/ objects is crazy. idg4 would have made a great edit model instead of just a t2i model.

u/Beautiful_Egg6188 — 2 months ago

Ideogram 4 on comfyui

High prompt adherence and control are the only reason to use it right now
Takes too long to generate
quality is decent but not as good as some other opensource models
Odd safety filter blocks on random.

u/Beautiful_Egg6188 — 3 months ago

LTX 2.3 Weird bug

there is this weird thing on the bottom of the screen that just doesnt go away. Ive tried generating multiple videos, with different resolution and settings. but this stays will all of them. how do i fix it

u/Beautiful_Egg6188 — 3 months ago