Skip to content

Author

Linjian Meng

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Aug 2026

Learning from Hard Prompts: Difficulty-aware Advantage Amplification in Dynamic Sampling

Direct Advantage Amplification (DAA), which amplifies the advantages of hard-to-sample correct responses on hard prompts, as obtained by Dynamic Sampling, is proposed, which ensures that, when Dynamic Sampling is used, these hard-to-sample responses can be effectively capitalized on, implying higher training efficiency.

Si-Yuan Gan, Yu-Hang Li, Xiran Wang et al. · 0 citations
#artificial intelligence Preprint Aug 2026

When Teacher Guidance Misleads: Reward-Aligned On-Policy Distillation

RA-OPD selects more reliable trajectories to improve student model performance without requiring additional computational cost and is evaluated on math and code benchmarks using models from the Qwen3 family and the DeepSeek-R1 family.

Si-Yuan Gan, Yu-Hang Li, Xiran Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.