Conference
Open access
Jul 2026
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation
The ViPo-MLLM model attained competitive performance compared to gloss-based recognition approaches, confirming the effectiveness of the proposed pose cues and cross-modal attention mechanisms.
A. Hasanaath, Bicheng Xu, Mir Rayat Imtiaz Hossain et al.
· International Conference on... · 0 citations