MiniMax H3 T2V/I2V
Took me a while to figure this one out, but I think this is a solid first version.
MiniMax is amazing, but a bitch to train for. Mainly because it learns too much and one or two items in your dataset can lead to weird behaviour.
For instance, too many closeups in this one makes it want to constanly zoom in or focus on the dick (which is kind of anoying when you don't want that). Ajusting the lora stregnth can help with this; you can get away with a as little as 0.6 strength if the subject is not in the forefront.
Keep in mind that the keyword is necessary for this one, otherwise you'll just get lumpy blobs.
I'll probably come back to this for a V2, but not until next month, Training is expensive as fuck, especially for video models.




