Skip to content

Author

Chengyin Xu

We have 2 of 2 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Scaling Domain Data Repetition in LLM Pretraining

This work finds that repetition counts tuned on smaller proxy models with the same \(\mathrm{TPP}\) can provide a practical estimate for larger models, and suggests that repetition counts tuned on smaller proxy models with the same \(\mathrm{TPP}\) can provide a practical estimate for larger models.

Jingwei Li, Xinran Gu, Rui Dai et al. · 1 citation
Preprint Aug 2026

psRL: Efficient Training for Agentic AI via Training-Time Prefix Sharing

This paper proposes psRL (prefix sharing for RL), a new training system for agentic AI designed to exploit prefix redundancy among training samples, and introduces two novel prefix-sharing mechanisms that enable flexible, fine-grained workload distribution across GPU workers.

Mian-Jie Yu, Zizhao Mo, Huanyu Qu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.