MV-Bench: Benchmarking Multimodal Large Language Models for Coordinated Multi-View Interface Construction
It is shown that current MLLMs can reproduce visual appearance but remain limited in generating the data semantics and interactive logic required by coordinated multi-view interfaces, and Iterative refinement improves code executability but does not substantially reduce the gap in data binding and interaction generation.