Jun 2026
IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations
IMCBench is introduced, an image-grounded, multi-turn medical conversation benchmark that pairs real, publicly available clinical images with synthetic patient profiles to simulate realistic patient-clinician interactions and demonstrates that accurate clinical description does not guarantee safe patient guidance, motivating the need for multi-dimensional evaluation frameworks in medical AI.
Maria Xenochristou, Ashutosh Joshi, Korosh Vatanparvar et al.
· arXiv.org · 0 citations