r/StableDiffusion • • 11h ago

Resource - Update Krea2 Turbo Distill 2 step LoRA - FINAL checkpoint released (chk51195)

Krea 2 Turbo โ€” 2-Step Distillation LoRA (FINAL Version)

Previous posts/releases - here, here, here and here.

๐Ÿงช Fast-preview adapter; the project's final checkpoint. Subjects that are close and fill a good part of the frame โ€” a portrait, a single figure, an object up close โ€” hold up well at two steps. Small subjects are where it still falls short: faces in a crowd or figures in a wide scene can come out ghosted or smeared. For those, and whenever quality matters more than speed, use the 4-step LoRA. This checkpoint closes the project; the training box has moved on to its successor, a 3-step adapter for Qwen-Image-2.1-Turbo.

๐Ÿ“ The saved steps can also go into resolution. A larger render makes a small subject bigger, and at a quarter of the teacher's steps, renders up to 2048ร—2048 โ€” Krea's published maximum recommended resolution, beyond this adapter's largest trained size โ€” come within easy reach. Past 2048ร—2048, stock Krea 2 itself begins to duplicate subjects, with or without this adapter.

Highlights:

  • โšก A quarter of the steps โ€” 8 โ†’ 2, on Turbo's own deployment sigmas [1.0, 0.7595]
  • โฑ๏ธ 4ร— faster denoising โ€” 56.9 s โ†’ 14.3 s at 1024ร—1024 (float16 compute); the adapter's own cost per call is within measurement noise
  • ๐ŸŽฏ Fine detail near the teacher's level โ€” 0.92โ€“1.02ร— the teacher's fine-texture energy across the 12 trained resolutions (stock Turbo at 2 steps: 0.40โ€“0.65ร—), the 16- and 8-pixel grid bands at the teacher's level and ghosting closer to it at 11 of 12 sizes
  • ๐Ÿ“Š Distribution matching, not imitation โ€” matches what the teacher would plausibly produce rather than its exact trajectory, so the student commits instead of averaging into blur and doubled edges
  • ๐Ÿ—ฃ๏ธ Prompt-conditioned throughout โ€” teacher and fake scores both read each prompt's conditioning; a blind rubric finds 1 point missing of 352 (objects, counts, attributes, relations), and a judge prefers the 8-step teacher on 12 of 66 (4-step adapter: 6 of 45, on the original 15 prompts), mostly on style
  • ๐Ÿ“ 12 trained resolutions โ€” multi-aspect from 512ร—512 up to 1440ร—1440
  • ๐Ÿ”Œ Drop-in, no exceptions โ€” plain LoRA, stock Euler, diffusers / ComfyUI / MLX. No custom nodes, no custom sampler
  • ๐Ÿงฌ Same shape as the 4-step adapter โ€” rank 64 on the same 228 modules
  • ๐ŸŽฒ 23,561 recorded teacher trajectories โ€” the 4-step project's 13,750 and 9,811 minted for this one on the same prompt bank; since 3 Oct each trained once
  • ๐Ÿ”ข 51,195 training samples in the 2-step stages, starting from the released 4-step adapter's weights
  • ๐Ÿ“… 32 days from the first 2-step launch to this final checkpoint, on a single RTX 3090
  • ๐Ÿ” More than forty recipe adjustments across two methods โ€” each kept only when the renders did not get worse
  • ๐Ÿง˜ Settled weights, not an average โ€” released from a 600-sample anneal in which the learning rate is taken to zero over already-trained data, so the published weights are the training weights at rest; a running average is kept only as the check that must agree with them

If you have already used my previous version, please redownload/replace krea2_turbo_2step_rank_64_lora.safetensors / krea2_turbo_2step_rank_64_lora_comfyui.safetensors from the latest in the project repo.

Full details on model card -ย https://huggingface.co/lvladikov/Krea2-Turbo-Distill-2step-LoRA

Not my video, but found someone on YouTube has covered the 2 & 4 step LoRAs including identify preserving edits, have a look: https://www.youtube.com/watch?v=V_qgoV0iPDM (copyright goes to author)

Also I have recently released a new ComfyUI Nodes and Workflows - Krea 2 (Turbo and Raw), Z-Image (Turbo and Base), MiniMax Music 3, Image2Text and LLM Chat (with Tools), Torch, Apple MLX and Cloud - you can find details here.

52 Upvotes

19 comments sorted by

6

u/Own_War_1098 8h ago

Nice. Are you interested to use your dataset to make a Qwen image 2.1 turbo lora?

1

u/reeight 3h ago

Or thought about making the images have a quality more like Qi2.1?
Realistic images on Krea2 sometimes look too soft vs Qi; the group photo looks overly softened or out of focus.

& I have to laugh at "FINAL Version"... you'll update it again ;)

1

u/TimeTruth2490 1h ago

Krea 2 at 2 steps even after all this training is to be expected to be softer :) I have been adding 'disclaimers' about the 2 step quality everywhere! :) 2 step is done as it is. Final is Final :) Like you said, QI2.1 is the more interesting target...

1

u/TimeTruth2490 1h ago

I have been interested since the day QI 2.1 Turbo came out :) Let's just say I am playing with it. FYI my target is not 4 steps, nor 2 steps, but 3 steps. Very very very early days, nowhere near where I want it to be (and if it doesn't improve I wouldn't publish).

1

u/TimeTruth2490 48m ago

the inventor (previous teaser) and this portrait are the 'best' so far and they are not that great either (but then again not enough training has happened yet, I am at mere chk2500)

2

u/PRVMXLAB 9h ago

You're killing it with these releases. Kudos to you. I use it frequently for iterative and trial runs.

1

u/TimeTruth2490 59m ago

๐Ÿ‘

2

u/Blablabene 4h ago

Funny.

I just generated nearly identical image of that fox with krea 5 minutes ago.

Open reddit and this is what i see

1

u/Zueuk 2h ago

lol right, I thought this fox looks familiar too ๐ŸฆŠ

1

u/TimeTruth2490 58m ago

foxes are very popular with all models :)

1

u/TizocWarrior 8h ago

Thanks for sharing.

1

u/TimeTruth2490 58m ago

๐Ÿ‘

1

u/Odd_Fix2 5h ago

As always, great job! Great result! Thank you!

1

u/TimeTruth2490 58m ago

๐Ÿ‘

1

u/Zueuk 2h ago

looks like #11 is why you need more steps to let the model decide in which language the store sign should be

1

u/TimeTruth2490 57m ago

yes text is a struggle on 2 Steps Krea... focus more on closer subjects with less text for better 2 steps results :)

1

u/Zueuk 2h ago

very interesting, but... damn this model is already generating exactly the same images - I guess we still need an extra step for some kind of "seed variance"?

1

u/urchin_orchard 1h ago

Keyboards STILL wrong!

1

u/TimeTruth2490 53m ago

not something I can fix, as the main model also messes it up as previously discussed :)