Train practically any diffusion image/video/audio model with SimpleTuner.
This model doesn't have a readme.