Skip to content

Author

Kaike Zhang

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Information-Time Proximal Policy Optimization

RLVR has substantially improved the reasoning capabilities of LLMs. However, existing methods typically parameterize temporal progression in the Markov Decision Process by token-by-token generation, despite the highly non-uniform information flow along autoregressive trajectories. In this paper, we propose InfoPPO, whi...

Yong-Cheng Zeng, Xin-Yu Cui, Yan Song et al. · 0 citations
Jul 2026

OpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party Skills

The behavioral analysis reveals three recurring failure patterns: agents may fail to recognize the risk, recognize it but fail to intervene before acting, or follow skill instructions beyond the user's intended scope, which highlights the need to improve both risk reasoning and execution control in agent frameworks.

Qi-Yuan Liu, Tingfeng Hui, Kun Zhan et al. · 3 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.