July 9, 2026 FSVideo: Fast Speed Video Diffusion Model in a Highly-Compressed Latent Space ByteDance's FSVideo achieves 42.3x speedup over Wan2.1 via 64×64×4 compression autoencoder, layer memory self-attention DIT, and few-step latent upsampler video-generation diffusion-model image-to-video latent-compression transformer