Preprint
Jul 2026
SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering
SeeMe is proposed, a training-free framework that introduces the concept of feature engineering from traditional machine learning into LVLMs and restructures visual tokens through a three-stage token engineering process to suppress hallucination sources while preserving informative visual evidence.
Kai Tang, Jinhao You, Bohua Zhang et al.
· 2 citations