VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding
This work introduces VideoChat3, a fully open, efficient, and generalist video-centric MLLM, which surpasses prior open-source models with equal or larger parameter counts with only 4B parameters and higher efficiency.