Skip to content

Author

Zijia Zhao

We have 2 of 15 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Jul 2026

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

PerceptionBench provides a capability-level standard for measuring and diagnosing the visual perception boundaries of MLLMs, by diagnosing the earliest failure points in the responses of frontier MLLMs across 42 existing benchmarks and constructing an error taxonomy whose perception branch defines ten atomic perceptual capabilities.

Zichao Lin, Yifeng Xie, Bowen Qu et al. · 2 citations
Jul 2026

TimeThink: Reasoning with Time for Video LLMs

TimeThink is proposed, a reinforcement learning framework that explicitly guides temporal evidence discovery in Video-LLMs and introduces a step-wise temporal process reward that provides localized credit assignment for these clues and a joint process--outcome optimization objective that balances reasoning fidelity with task correctness.

Handong Li, Longteng Guo, Zikang Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.