Malignant skin lesions are often visually indistinguishable from benign analogs, presenting a challenging "cancer-mimic" scenario. Post-hoc methods, such as Grad-CAM and SHAP, fail to ground model decisions in the ABCD clinical criteria (Asymmetry, Border, Color, Diameter) relied upon by dermatologists. To address this...
Minh Nguyen, Hien Thi Thuy Phan, T. Nguyen et al.· International Conference on...· 0 citations
Large language models (LLMs) have demonstrated strong capabilities across diverse domains, showing considerable potential in medicine. However, their application in medical settings remains limited by the scarcity of visual question answering (VQA) datasets that capture clinical reasoning and explicit image-text alignm...
Ling-Xuan Hou, Yu-Hua Xie, Yue Hu et al.· 0 citations
A FOundational LLM Trained on ThoughtMed-1M (FOLTMed), a scalable paradigm for advancing research on clinically grounded multimodal LLMs, achieved state-of-the-art performance across 42 medical VQA benchmark datasets, with a macro accuracy of 85.4%, and generated more clinically coherent responses on the ThoughtMed-1M...
Ling-Xuan Hou, Yu-Hua Xie, Yue Hu et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.