Video Generation Models Are Inherent Lighting Estimators
V-LITE (Video generation models are inherent lighting estimators), a framework that unlocks internal knowledge by reframing lighting estimation as a guided video inpainting task, is introduced, revealing that modern video diffusion models are not merely synthesizers but also powerful, inherently capable estimators of physical scene lighting.