ProCap: Prominence-guided Object Rectification for Faithful and Comprehensive Video Captioning
A prominence-aware, iterative post-hoc rectification framework that overcomes both limitations without modifying the underlying captioning model's parameters, and position prominence-guided iterative rectification as a lightweight, scalable, and model-agnostic route to more complete and trustworthy video captioning.