From Single- to Cross-Document: Benchmarking Multi-Granularity Event Analysis of Large Language Models
MGUE-Bench is introduced, a systematic benchmark for assessing the performance of large language models in multi-granularity event analysis, and extensive experiments on state-of-the-art LLMs and retrieval-augmented generation methods delineate the current capability boundary and identify critical deficiencies, providing insights into the future improvement of LLMs in challenging event analysis tasks.