Skip to content

Author

Rui-Jia Chen

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Oct 2026

Flash-OPD: Fast On-Policy Distillation

On-policy distillation (OPD) provides dense teacher supervision on student-generated trajectories, but generating and evaluating long rollouts incurs substantial training cost. Existing acceleration methods reduce this cost through open-loop rollout schedules or closed-loop horizon adaptation. However, supervision comp...

Wei Chen, Junle Chen, Yi-Tong Yang et al. · 0 citations
#artificial intelligence Preprint Oct 2026

You Changed Your Mind, The Model Didn't: Demystifying Intent in Multi-Turn Dialogue

When a large language model handles a multi-turn task and a user proposes a change but ultimately rejects it, the model should continue as if nothing changed. We find a surprising failure: merely mentioning a rejected change can derail task execution, even when the user's final intent remains unchanged. To systematical...

Junle Chen, Wei Chen, Zheng-Jun Huang et al. · 0 citations
#natural language process... Preprint Oct 2026

My FAULT: Self-Diagnosis as Credit Assignment in Self-Evolving Agentic Reinforcement Learning

Agentic reinforcement learning (RL) has emerged as a powerful approach for training large language model agents on multi-step tasks, yet reliance on terminal outcome rewards creates two credit-assignment problems, particularly in long-horizon tasks. First, same-outcome rollout groups provide no learning signal from ter...

Yi-Hua Zhu, Qian-Ying Liu, Wei Qiao et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.