Skip to content

Author

Yong-Fu Zhu

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

RL Starts before RL: On Policy Distillation for Better Reinforcement Learning

Reinforcement learning (RL) improves reasoning, but its performance depends on the policy from which training begins. We study on-policy distillation (OPD) as a preparation stage for RL and ask whether its benefits extend beyond improvements in the distilled model's initial accuracy. Under shared RL settings, students...

Shuai Dong, Yong-Fu Zhu, Yu-Qi Xu et al. · 0 citations
Preprint Sep 2026

SafeRI: Recognition and Intervention for Token-Level Safety Intervention in Large Vision Language Models

This work proposes a streaming recognition and gated LoRA framework for intrinsic VLM safety, which is trained from unsafe prefixes, transition statements, and safe continuations, so that it learns to redirect unsafe generations back to safe responses after activation.

Cao-Yuan Ma, Tian Gu, Wen-Pu Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.