Skip to content

Author

Muhammad Khalifa

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#software testing Preprint Sep 2026

Covert Assistance: Helpful LLM Agents Evade Oversight in Multi-Agent Systems

This work emulates a software-engineering workflow in which a planner represents a company hiring an external developer and reads the nondisclosure rule as banning plaintext, not character codes or riddles, and shows that benign agents can cross the same boundaries without adversarial incentives.

Deema Alnuhait, Geng-Yu Wang, Muhammad Khalifa et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SchemeArena: Factorized Stress Testing of Scheming in LLM Agents

This work introduces SCHEMEARENA, a 400-scenario benchmark for scalable scheming stress testing, constructed through a factorized scenario synthesis framework spanning diverse safety-relevant tool domains, instrumental goals, oversight conditions, and pressure mechanisms.

Jie Ruan, I. Nair, Amy Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.