Skip to content

Author

Zhan Qin

We have 3 of 11 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Review Open access 2026

A survey on AI agent security: Reasoning, Acting, and Self-Evolving

Artificial intelligence (AI) agents are rapidly evolving into autonomous systems capable of reasoning, acting, and continual evolving. Their increasing autonomy enables powerful real-world applications but also introduces security risks throughout the entire agent execution cycle. However, the security challenges arisi...

Dai-Zong Liu, Tian-Yao Luo, Shu-Wei Huang et al. · 0 citations
Jul 2026

DARWIN: Evolving Jailbreak Adversary and Guardrail for LLM Safety Evaluation and Protection

DARWIN, an evolutionary attack-defense framework that models jailbreaking as a continual process and updates guardrails through an attack-defense loop is proposed, an evolutionary attack-defense framework that models jailbreaking as a continual process and updates guardrails through an attack-defense loop.

Weiwei Qi, Ze-Feng Wu, Zhiling Guo et al. · 4 citations
Preprint Jul 2026

DataShield: Uncovering Risky Fine-Tuning Data Across LLMs Through Consensus Subspace Alignment

DataShield is a data assessment framework that identifies risky fine-tuning samples and response segments through consensus subspace alignment over joint safety-critical semantic spaces derived from multiple safety-aligned LLMs, allowing both sample-level filtering and fine-grained segment-level masking.

Ze-Feng Wu, Weiwei Qi, Jielong Chen et al. · 4 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.