Jul 2026
Group-of-Latents: Perceptual Video Compression at Extreme Bitrates via Masked Latent Generative Modeling
A unified generative framework that leverages pre-trained Diffusion Transformer priors to achieve high perceptual quality at extremely low bitrates, achieving state-of-the-art perceptual fidelity with rich spatial details and robust temporal consistency.
Shaokang Wang, Jinchang Xu, Peidong Jia et al.
· arXiv.org · 0 citations