When does training on downscaled images yield the same gradients?
Paper • 2608.04448 • Published
Anima lora for turbo generation, implemented with DP-DMD and Continuous-Time Distillation paper.
This model is presented in the paper When does training on downscaled images yield the same gradients?.
Try new anima_turbo_4step_v3.safetensors
trained with 10k steps, 30h in single 5070ti using anima_lora
Reproducible by CFG=1.0, steps=4, er_sde sampler
Base model
nvidia/Cosmos-Predict2-2B-Text2Image