Skip to content

Category

cybersecurity

1,065 papers

FAA Framework: A Large Language Model-Based Approach for Credit Card Fraud Investigations

The first Fraud Investigation Assistant (FIA) framework is introduced, which employs multimodal large language models (LLMs) to automate key steps of credit card fraud investigation and generate explanatory reports and suggests that LLM-based agents can assist with automating substantial parts of the fraud investigatio...

Shaun Shuster, Eyal Zaloof, A. Shabtai et al. · 5 citations · ⚡1

Hidden Thoughts Are Not Secret: Reasoning Trace Exposure in LLMs

REP, a lightweight in-context elicitation method that uses shadow-model-generated demonstrations wrapped in auxiliary code-like formats to raise user-visible reasoning traces from a victim model, substantially increases similarity between exposed and REP-conditioned internal traces while preserving useful reasoning sig...

Yu-An Lu, Ci-Yang Tsai, Yu-Lin Tsai et al. · 1 citation · ⚡1
#artificial intelligence Preprint Aug 2026

SIR: Self-improving Red-teaming for Compute Use Agents

SIR is presented, a black box IPI attack that composes stealthy injections from a small library of reusable principles stated in plain language and wraps composition in an iterative feedback loop that diagnoses the victim's failed trajectories and distills the bypasses into new, named strategies that are reapplied acro...

Chen Xiong, Zhi-Yuan He, Pin-Yu Chen et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Influence Is Not Authority: When Causal Guardrail Signals Make Legitimate Tool Use Look Like an Attack in Tool-Using LLM Agents

The results show that the studied causal signal reveals what shaped an action without reliably encoding whether the action was authorized, and that reference construction and routing are integral to the effective security decision.

Tanzim Ahad, Ismail Hossain, Md. Jahangir Alam et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Safe to Resume? Breaking Execution Continuity of Agent Execution via Rollback

This paper presents the first systematic security study of checkpoint and rollback in existing agent systems, and develops a multi-agent analysis pipeline that reconstructs execution semantics, identifies violations of the five failure conditions, and validates them through actual rollback.

Guan-Long Wu, Da-Hui Li, Ke Jiang et al. · 2 citations
#artificial intelligence Preprint Aug 2026

Auditing and Mitigating Privacy Leakage in Cloud-Edge Collaborative Decoding

CoVeil is proposed, a defense mechanism which dynamically optimizes transmitted signals to suppress leakage during decoding time while preserving the collaborative quality, and consistently improves the privacy-utility trade-off over existing baselines by reducing data leakage.

Ke-Jia Zhang, Tianyuan Zou, Zi-Xuan Gu et al. · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.