Skip to content

Author

Min Zhang

We have 3 of 19 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Jul 2026

Agent Reinforcement Learning via Pivotal-Aware Self-Feedback Retry

PivoARL is proposed, a self-feedback retry framework for experience exploitation in LLM agents that identifies the pivotal erroneous turn through structured reflection and performs local retry only from the corresponding pivotal state, thereby reusing the correct prefix and reducing redundant interactions.

Weiyang Guo, Zesheng Shi, Longhui Zhang et al. · 2 citations
Preprint Aug 2026

From Profiling to Synthesis: Benchmarking Implicit Behavioral Alignment in Personalized LLM Agents

IBA-Bench is introduced, a benchmark for implicit behavioral alignment constructed from longitudinal interaction histories that contain noise, implicit cues, and temporal inconsistencies, and the proposed IBA-Agent is proposed, an agent framework that reconciles conflicting priorities through broad retrieval and trajectory-level alignment.

Jiajia Song, Bobo Li, Haiwen Yi et al. · 0 citations
#artificial intelligence Preprint Jul 2026

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

This work demonstrates a full-stack pathway from efficient trillion-parameter model post-training on Ascend infra to domain-specialized Flash models for solver-grounded mathematical modeling, advancing frontier-model systems for complex reasoning.

Dongfang Li, Xiaodong Luo, Ruoyu Sun et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.