Skip to content

Author

Ran He

We have 2 of 15 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

When Not to Imitate: Boundary-Aware Skill Memory for Reliable Tool-Use LLM Agents

BASM is proposed, which augments each skill with explicit boundary fields, which transforms each retrieved skill from an unconditional action template into state-conditioned guidance: the agent applies the skill when its conditions hold, suppresses inapplicable tool calls when they do not, and issues targeted repairs when execution fails.

Zi-Han Lin, Zhenyu Chen, Jiawen Wei et al. · 0 citations
Jul 2026

TREK: Distill to Explore, Reinforce to Refine

TREK (Teacher-Routed Exploration via Forward KL), a simple staged procedure that uses distillation not for imitation but for exploration support expansion, achieves high success rates early in training while unaided GRPO requires substantially more optimization steps to reach comparable levels.

Yuanda Xu, Zhengze Zhou, Kayhan Behdin et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.