VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video ModelsHila Chefer, Shelly Sheynin, Yuval Kirstain et al.|arXiv (Cornell University)|2025Cited by 0