Skip to content

Author

D. Song

We have 3 of 13 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

AgentXploit: Autonomous Repository-to-Runtime Red-Teaming for AI Agents

This work presents AgentXploit, a two-role auditing system that separates repository-level attack-path discovery from runtime exploitation and introduces AgentXploit-Bench, containing 72 reproducible vulnerabilities across 12 open-source AI-agent systems and frameworks.

Wei-Da Liang, Shi Qiu, Zhun Wang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

BenchShield: Formal Model-Backed Instrumentation for Reward Integrity in LLM-Agent Evaluation Infrastructure

BenchShield is presented, a model-backed instrumentation layer for reward integrity in LLM-agent evaluation that grounds detection in a finite lifecycle model of an evaluation's reward-relevant events and achieves 96% accuracy in detecting reward hacking from infrastructure-side evidence.

Sheng-Han Zheng, Zong-Lin Di, Yimin Liu et al. · 0 citations
Jun 2026

Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

SafeClawArena is developed, a benchmark of 406 adversarial tasks executed in containerized replicas of real agent platforms with canary-marked credentials and evaluated via automated taint tracking across nine output channels, exposing the inadequacy of current defenses and suggesting directions for future hardening.

Peizhi Niu, Wenjie Qu, Shangding Gu et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.