Improved Video VAE for Latent Video Diffusion Model
ByteDance/Tongyi Lab's IV-VAE introduces Keyframe-based Temporal Compression and Group Causal Convolution to achieve SOTA video reconstruction with balanced inter-frame performance
1 post tagged with "causal-convolution"