Skip to content

Author

Yi-Feng He

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

SRE-Marathon: A Continuous, Change-Driven Benchmark for Autonomous Site Reliability Agents

This work presents SRE-Marathon, a benchmark for long-horizon, continuous SRE operation, where an agent is invoked at a fixed cadence with cumulative alert history and a persistent workspace while operating a live two-zone Kubernetes deployment as a fault orchestrator injects overlapping faults according to a seeded, p...

Yi-Fang Tian, Ying-Jian Bai, Yi-Feng He et al. · 0 citations
#artificial intelligence Preprint Sep 2026

BenchShield: Formal Model-Backed Instrumentation for Reward Integrity in LLM-Agent Evaluation Infrastructure

BenchShield is presented, a model-backed instrumentation layer for reward integrity in LLM-agent evaluation that grounds detection in a finite lifecycle model of an evaluation's reward-relevant events and achieves 96% accuracy in detecting reward hacking from infrastructure-side evidence.

Sheng-Han Zheng, Zong-Lin Di, Yimin Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.