Threat hunting increasingly depends on converting unstructured knowledge (e.g., Cyber Threat Intelligence reports) into actionable hunt leads: concise, investigable hypotheses grounded in observable artifacts and adversary techniques. Producing such leads manually is a tedious and hard-to-scale task. Existing automated...
A. Prakash, Boubakr Nour, M. Pourzandi et al.· 0 citations
Large language model (LLM) benchmarks are often treated as fixed datasets with stable scores, yet their outcomes depend on configurable evaluation pipelines. We audit eight cybersecurity benchmarks across 10 proprietary, open-weight, and cybersecurity-specialized LLMs. By modeling benchmarks as measurement pipelines, w...
Aymene Berriche, Cathrine Shalby, Mohannad J. Alhanahnah et al.· 1 citation
A platform that connects a pluggable red-team adapter and a pluggable blue-team adapter to a shared target LLM and scores their attack and defense rates with an LLM judge, and describes the design of ACEA and the metrics through which red and blue teams are scored head to head.
Yi-Da Shen, Kentaroh Toyoda, Alex Leung· 0 citations
A controlled study on six open-source research software projects, with protocol, seed, panel, and analysis plan deposited with a DOI before any trial, drew three conclusions: publishing signals is necessary but not sufficient; price did not buy verification; verification must be built into the program that runs the ass...
Pengyin Shan· 0 citations
Reach audiences
Advertise in front of researchers, engineers, and readers.
Repeated runs show that category-level and within-trajectory relations can recur even when normalized score-change rankings do not, and motivate agent-behavior evaluation that links communication, authorization, and evolving state instead of treating individual transaction verdicts as complete safety judgments.
Ze-Lin Li, Yi-Yun Su, Matt White et al.· 0 citations
Intentest is proposed, an intent-graph-guided automated penetration testing agent that externalizes long-horizon state from the LLM's context window onto a persistent fact-intent directed acyclic graph (DAG), thereby substantially reducing invalid transitions.
Wei-Zhe Wang, Yi-Tong Zhang, Yao Zhang et al.· 0 citations
It is established that adversarial robustness in ICS anomaly detection is a specific instantiation of system resilience, and a compositional resilience bound for heterogeneous ICS detection networks is derived, showing that the binding constraint on system-level resilience is the coupling-adjusted absorption capacity o...
Branka Stojanović, Andreas Flatscher, Michael Somma· 0 citations
The findings reveal a confidentiality risk in LLM agents: protecting explicit artifacts alone is insufficient, as observable execution behavior can leak the procedural knowledge required to reconstruct proprietary task-solving capabilities in low-capability and attacker-controlled agents.
Xiaoting Lyu, Yu-Hong Wu, Yu-Fei Han et al.· 1 citation
This work introduces CIPHER (Cross-record Inference over Privacy-Hardened Evidence Records), a benchmark of expert-validated questions from consumer-finance, clinical, and law-enforcement records that evaluate retrieval, prompting, table-specialist, and hybrid symbolic-neural systems under native redaction and surrogat...
Suparno Roy Chowdhury, M. Choudhury, Dhruv Madhwal et al.· 0 citations
AgentDrift is presented, a benchmark of 12,536 synthetic tool-call trajectories over five agent domains in which every one of the 71,024 steps carries one of four labels: benign, injection point, hijacked, or failed injection; it is shown that the LLM judge was itself fooled by the hard negatives.
Asif Pinjari, Mithun Paul Saint-Germain· 2 citations
Overall, KV-cache timing reliability depends strongly on the load regime and serving stack, and measurements on quiet systems can overestimate operational attack reliability.
Web Application Firewalls (WAFs) mainly rely on signatures to detect known attacks, which can leave gaps against modified or previously unseen payloads. Positive security provides a complementary approach by learning legitimate traffic and blocking inputs that fall outside the learned profile. However, learning directl...
H. Osama, Zeyad Ahmed, M. Amgad et al.· 0 citations