Jul 2026
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding
This work introduces VideoChat3, a fully open, efficient, and generalist video-centric MLLM, which surpasses prior open-source models with equal or larger parameter counts with only 4B parameters and higher efficiency.
Xinhao Li, Yuhan Zhu, Xiangyun Zeng et al.
· arXiv.org · 4 citations