Author

Donghan Guan

1 paper indexed here

Fetches their full publication history.

Not the right person? Other researchers publish under this name.

Retrieval-Augmented Multimodal Large Language Models for Visual Question Answering of Construction Occupational Health and Safety Hazards

A visual knowledge enhancement framework for construction OHS visual question answering (VQA) based on multimodal large language models (MLLMs) and retrieval-augmented generation (RAG) is proposed, augmenting managerial capacity for reliable and objective OHS hazard prevention.

Yang Liu, Luping Li, Xing Su et al. · 0 citations