If the LoRA you're after is of a specific looking (or just specific) person, just my opinion, but I would opt for working on the LoRA with an image model. That way I could use it to place that person in any environment I wanted via prompts, and then send the images through a video model like Wan or LTX. I think it is harder and more hardware intense to try to train a video LoRA for a specific person.
Am I understanding that to be your goal? Train a video LoRA to produce a specific subject?
Am I understanding that to be your goal? Train a video LoRA to produce a specific subject?
