RGB-T Salient Object Detection (RGB-T SOD) effectively leverages the complementary information of RGB images and thermal infrared images to locate important targets in complex environments, such as low light, rainy and foggy weather, or cluttered backgrounds. However, the existing deep-learning based models have two ke...
Ze Li, Ying-Ying Zhang, Shuai Zhang et al.· Neural Networks· 0 citations
Large language models (LLMs) have demonstrated strong capabilities across diverse domains, showing considerable potential in medicine. However, their application in medical settings remains limited by the scarcity of visual question answering (VQA) datasets that capture clinical reasoning and explicit image-text alignm...
Ling-Xuan Hou, Yu-Hua Xie, Yue Hu et al.· 0 citations
A FOundational LLM Trained on ThoughtMed-1M (FOLTMed), a scalable paradigm for advancing research on clinically grounded multimodal LLMs, achieved state-of-the-art performance across 42 medical VQA benchmark datasets, with a macro accuracy of 85.4%, and generated more clinically coherent responses on the ThoughtMed-1M...
Ling-Xuan Hou, Yu-Hua Xie, Yue Hu et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.