Skip to content

Author

Hao Yang

We have 2 of 19 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Jul 2026

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

PerceptionBench provides a capability-level standard for measuring and diagnosing the visual perception boundaries of MLLMs, by diagnosing the earliest failure points in the responses of frontier MLLMs across 42 existing benchmarks and constructing an error taxonomy whose perception branch defines ten atomic perceptual capabilities.

Zichao Lin, Yifeng Xie, Bowen Qu et al. · 2 citations
Jul 2026

Towards Predictive, Aligned, and Scalable Robot Learning

Lumo-2 is introduced, a latent world-action model that generates actions by reasoning over world dynamics in latent space that consistently outperforms strong vision-language-action and world-action model baselines, with gains on challenging real-world tasks requiring temporal reasoning, physical understanding, or high control complexity.

Peijun Tang, Shang-Ping Xie, Baifu Huang et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.