Skip to content

Author

Yi-Ming Zong

We have 3 of 9 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

CERO: Where and When to Allocate Rollouts for RL Post-Training

Adaptive rollout methods for group-relative reinforcement learning typically allocate a fixed per-update budget across prompts. We instead study how to coordinate a finite rollout budget over the entire training horizon. We formulate this problem using a concave surrogate utility of cumulative prompt exposure and intro...

Yi-Ming Zong, Yi-Ge Wang, Xin-Ting Hu et al. · 0 citations

Adaptive Resolving Methods for Markov Decision Processes with Function Approximations

This work considers the MDP problems with function approximation with function approximation and develops a new algorithm to solve it efficiently, based on a linear programming (LP) reformulation and repeatedly resolves the identified reduced linear system as new transition samples arrive.

Jia-Shuo Jiang, Yinyu Ye, Yiming Zong · 0 citations
#artificial intelligence Preprint Aug 2026

Learning to Allocate Incentives for Incentivized Advertising via Offline Model-Based Reinforcement Learning

An offline model-based RL framework for cost-controllable sequential incentive allocation is developed and an independent counterfactual scorer evaluates each learned policy on held-out logs, enabling pre-launch selection without costly online exposure.

Zi-Lin Zhao, Han Yang, Tian-Pei Yang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.