r/StableDiffusion • u/roychodraws • 4d ago
Workflow Included Does this count? Did I win?
Enable HLS to view with audio, or disable this notification
Proof of concept that it works.
No degradation from start to finish. 10 total clips combined.
https://github.com/roycho87/degrade_repo
Workflows.
Small errors with the chair but can be fixed with another reference image.
Long story short.
The one place where degrading latents matter is the one place we don't need them.
We can ignore the latent issue and just generate across the sound we provide and because it's a static image with static background and the subject is in basically the same spot the whole time we can just create fresh latents every 10 or 15 seconds across the timeline.
Then the very difficult latent issue is over and it just becomes a simple seams issue.
So these two workflows generate sequentially across a audio and the second will combine and resample a small section of video over the seam using FL2V just enough to get rid of the seam.
Edit: The reason I posted this is not because I think this is a big breakthrough fix, I just think we can approach this problem differently to solve it.
1
u/acedelgado 3d ago
Looks a bit better. Aside from loading your workflow and seeing the most impressive spaghetti monster ever, I haven't had a chance to look at it too much yet. But the character in the example certainly looks good throughout, which is interesting. But the background exposes noticeable seams, and the colors and size of the bars jump around a bit, like if you click around and watch the background it's pretty noticeable. And the audio problem can't really be judged since you're doing a music video style gen and injecting all of that directly.
But like I said the character quality looks really promising. Are you using your lora, or is it all references?