Skip to content

Author

Zheng Zhang

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

Choosing Before Acting: Comparative Value Estimation for Long-Horizon Tool-Use Agents

Large language models (LLMs) rely on long-horizon tool invocation sequences for complex tasks, where each invocation can alter the task state and condition subsequent decisions. In long-horizon tool use, final-outcome rewards provide weak credit assignment over long interaction traces. Step-level rewards can offer more...

Yu Li, Zheng Zhang, Xin Liu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Rep2Skill: Representation-Guided Skill Self-Evolution for LLM Agents

Textual skills enable large language model (LLM) based agents to accumulate reusable procedural knowledge without updating model parameters. Yet existing skill evolution remains largely confined to the text space: an optimizer must diagnose success and failure patterns, and revise skills solely from long execution traj...

Kai-Xin Zhang, Chang-Ming Li, Ying-Dong Shi et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Train Ahead, Distill Back: Bootstrapping On-Policy Self-Distillation for Large Language Models

On-policy self-distillation (OPSD) improves large language models by letting a self-teacher with privileged information provide dense token-level supervision on the model's own trajectories. Yet existing methods typically construct the self-teacher from the current, initial, or slowly averaged policy state, leaving the...

Zheng Zhang, Xin-Yue Tan, Lu Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.