V3 update: Included smash cut type transitions. "The camera makes a smash cut..."
As much as I love Wan, if it's one thing it's not good at, it's cuts.
This lora is an experiment in greater shot control, by being trained on hard cuts. Most of the clips are also cut on action, fwiw, to allow for a continuous flow.
It's captioned with "wide-angle", "mid-shot" and "close-up". But I think Wan already handles them well.
Prompt format:
[a brief description of your initial shot.]
the camera makes a hard cut to [resulting shot]
[what happens after]
You can string several cuts. If you do, it works best if you also change the type of shot.
Where's the low noise model? There isn't one. It seems to work well with only the high noise model. I'll still train and test a low noise model, to see if it will improve consistency during cuts.
Play with strength if the scene changes too much. Too low strength will make it a dissolve transition (at least testing seems to imply that).
Using lora's (for the target cut) you can get pretty far.
If the cut doesn't 'take':
Try using less other lora's.
If you use a high resolution, lower it.
Try generating a shorter clip.
Description
Larger dataset, more variety.
FAQ
Comments (12)
The Lora is Awesome ! Thanks
This is a very useful LoRA! thanks for putting it together.
I'd like to ask, what form of data is used to train this LoRa? Is it a two-stage edit where one shot switches to another, with corresponding storyboard prompts?
Approximately what order of magnitude and number of training steps are needed for it to be effective?
Yes, the format described in the info.
@edwarqzheng6589 v2 is about 1800 steps.
Furthermore, I noticed that with some non-realistic image inputs, LoRa tends to render the images more realistic. Could this be because the training data distribution is more biased towards a cinematic style?
It could be.
Your examples don't have example prompts. Can I assume that the prompts are the same format as earlier versions, except "hard cut" is replaced with "the camera makes a hard cut"?
The mp4s should have the workflow embedded. But yes, replace 'hard' with 'smash'. However, I think it makes the distinction based on the descriptions.
I misread your comment. It's always been "the camera makes a hard cut". But I guess I had been lazy when adding the trigger.
Significant upgrade in terms of prompt coherence over the Deepfake lora I was using before this. Not to mention the much faster and hard cut, as promised.
Details
Files
hard_cut_2_100_wan_i2v_high.safetensors
Mirrors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
HardCut_HIGH.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
hard_cut_2_100_wan_i2v_high.safetensors
HardCut_HIGH.safetensors
Available On (1 platform)
Same model published on other platforms. May have additional downloads or version variants.