Skip to content

Author

Huan Liu

We have 2 of 18 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Measuring and Detecting Harmful AI Sycophancy

It is demonstrated that detection performance drops on unseen models and an initial approach is proposed to address this challenge, and it is shown that detecting PSRS is feasible from the response text alone, and detectors need to learn subtle PSRS patterns from the training data.

Bohan Jiang, Dawei Li, Yasin N. Silva et al. · 0 citations
Preprint Jul 2026

TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text

Comprehensive experiments demonstrate that TextCloak consistently impairs unauthorized fine-tuning while maintaining text utility for legitimate use, highlighting its broad applicability as a practical defense against unauthorized LLM exploitation.

Chengshuai Zhao, Pingchuan Ma, Dawei Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.