Can multimodal large language models evaluate AI-generated industrial design images? An empirical comparison with human expert assessment
Text-to-image (T2I) generation models are increasingly used in industrial design, but assessing the quality of their outputs still depends on slow, small-scale human expert review, which constrains the automation advantage of generative AI and hinders its adoption in this domain. This paper empirically examines whether...