TY - JOUR TI - Align Your Latents: High-Resolution Video Synthesis with Latent Diffusion Models AU - Blattmann, A. AU - Rombach, R. AU - Ling, H. AU - Dockhorn, T. AU - Kim, S. W. AU - Fidler, S. AU - Kreis, K. PY - 2023 DA - 2023/04// JO - CVPR 2023 (IEEE/CVF) AB - Blattmann et al. (2023) extend the latent diffusion model architecture to video by inserting temporal attention and 3D convolution layers into a pretrained image LDM, aligning the temporal dimension to produce temporally coherent high-resolution video. Published at CVPR 2023, the paper is the architectural ancestor of the latent video generation category. Admitted as a distinct foundational entry per owner decision QL3, it provides the conceptual origin for video-generation tools, including AnimateDiff, that are entering animation and motion-graphics production workflows. KW - generative-ai KW - video-generation KW - production-practice UR - https://arxiv.org/abs/2304.08818 LA - en ER -