Skip to content

Author

Zihan Zhang

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Optimal Multi-Reward Reinforcement Learning

We study an unknown-transition finite-horizon Markov decision process (MDP) with a finite collection of known reward functions $\{r^1, r^2, \ldots, r^M\}$. The goal is to output an $\epsilon$-optimal policy for every reward using online episodic interaction only. Performance is measured by the policy error $V_{0}^{*, m...

Zi-Jun Chen, Zi-Han Zhang · 0 citations
Jul 2026

Asymptotically Optimal Regret for Reinforcement Learning without Horizon Dependence

A new bound on total deviation for time-homogeneous MDPs is proved and a cutting bonus that preserves both optimism and the monotonicity needed for planning is designed that preserves both optimism and the monotonicity needed for planning.

Runlong Zhou, Zi-Han Zhang, Maryam Fazel et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.