Skip to content

Category

cybersecurity

1,065 papers

#artificial intelligence Preprint Open access Sep 2026

Hearing the Unspoken: Language Model Priors for Acoustic Adversarial Attacks

Automatic Speech Recognition (ASR) systems operating in real-time settings must process acoustic input under strict temporal constraints, where transcription decisions are inherently made on incomplete information. This causal constraint serves as an information bottleneck on attackers, significantly limiting attack pe...

Jiani Xie, Andrew C. Cullen, Paul Montague et al. · 0 citations
#artificial intelligence Review May 2026

IntraGuard: Committee-Side Defenses Against Review Outsourcing to Commercial Chatbots

IntraGuard is proposed, a black-box, venue-agnostic defense framework grounded in the structural--visual decoupling inherent to the PDF that achieves a defense success rate of up to 84%, while preserving peer-review invariance for human reviewers.

Ou-Bo Ma, Rui-Xiao Lin, Jia-Hao Chen et al. · 2 citations

An AI Agent Execution Environment to Safeguard User Data

This paper presents GAAP (Guaranteed Accounting for Agent Privacy), an execution environment for AI agents that guarantees confidentiality for private user data deterministically, without trusting the agent with private user data, and without requiring any AI model or the user prompt to be free of attacks.

Robert Stanley, Avirishu Verma, Lillian Tsai et al. · 4 citations
#artificial intelligence Preprint Open access Sep 2026

Measuring and Exploiting Contextual Bias in LLM-Assisted Security Code Review

Automated Code Review (ACR) systems integrating Large Language Models (LLMs) are increasingly adopted in software development workflows, ranging from interactive assistants to autonomous agents in CI/CD pipelines. In this paper, we study how LLM-based vulnerability detection in ACR is affected by the framing effect: th...

Dimitris Mitropoulos, Nikolaos Alexopoulos, Georgios Alexopoulos et al. · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Can We Stop Malicious AI? KILLBENCH: A Benchmark for External AI Kill Switch Feasibility

Malicious AI causing harm to humans is not just a Hollywood fantasy. Indeed, as highly capable models such as Claude Mythos emerge and agent systems like OpenClaw rapidly spread, the question of how to stop an AI that acts maliciously -- whether by design or by accident -- has become urgent. To address this, we propose...

Sechan Lee, Hyounghun Kim, Sangdon Park · 0 citations
#artificial intelligence Preprint Open access Sep 2026

Data Security in Large Language Models: Risks, Defense, and Directions

Large Language Models (LLMs), now a foundation in advancing natural language processing, power applications such as text generation, machine translation, and conversational systems. Despite their transformative potential, these models inherently rely on massive amounts of training data, often collected from diverse and...

Kang Chen, Xiuze Zhou, Yuanhui Yu et al. · 0 citations
#artificial intelligence Preprint Open access Sep 2026

SynGhost: Invisible and Universal Task-agnostic Backdoor Attack via Syntactic Transfer

Although pre-training achieves remarkable performance, it suffers from task-agnostic backdoor attacks due to vulnerabilities in data and training mechanisms. These attacks can transfer backdoors to various downstream tasks. In this paper, we introduce $\mathtt{maxEntropy}$, an entropy-based poisoning filter that mitiga...

Pengzhou Cheng, Wei Du, Zongru Wu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Vulnerability Localization Benchmark: Measuring Agentic Security Analysis at Repository Scale

VLoc Benchmark results establish vulnerability localization as a distinct repository-scale capability and provide a setting for studying both how security agents search for vulnerable code and when they should refrain from reporting it.

Aman Priyanshu, Supriti Vijay, Kimia Majd et al. · 0 citations
#artificial intelligence Review Sep 2026

LLM-Based Schema-Aware Split Learning for Privacy-Preserving Mental Distress Prediction Across Heterogeneous Surveys

Rising societal and lifestyle complexity has been linked to a growing prevalence of mental distress worldwide. Educational institutions, workplaces, clinics, etc. collect large volumes of mental health survey data to understand and reduce this burden. Collaborative analysis of such data could yield effective generaliza...

Md. Khalid Syfullah, Alvi Ataur Khalil · 0 citations
#artificial intelligence Preprint Sep 2026

When the World Lies: Backdoor Attacks on Latent World Models for Downstream Control

Pretrained world models, learned simulators that encode an observation into a latent state and predict how it evolves under actions, are beginning to be reused as off-the-shelf dynamics backbones for control, like pretrained encoders and language models are reused today. We show that this reuse opens a supply-chain bac...

Roberto Riaño, Gorka Abad, S. Picek et al. · 0 citations

From tech blogs

See all →
Google DeepMind Blog Jul 17, 2026

Introducing Gemini 3.5 Flash Cyber

Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model to find and patch vulnerabilities.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.