Skip to content

Author

Samuele Poppi

We have 5 of 19 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Oct 2026

Certification of Real Images through Calibrated Content Authentication

Generative models can synthesize high-quality inauthentic multimedia content that is already being misused at scale. We evaluate twenty deepfake detectors against ten generators released in the last four years and find accuracy decreasing over time, from near-perfect 99.5% to 76%. Adversarial perturbations further redu...

Sarim Hashmi, Abdelrahman W. A. Elsayed, Mohammed Talha Alam et al. · 0 citations
#machine learning Preprint Aug 2026

Context Inference Attacks Without Jailbreaks

This work introduces and formalizes context-inference attacks through a security game and evaluates three settings under decreasing attacker knowledge and increasingly indirect delivery of the context: a known context, an unknown context, and a context the agent retrieves through its own tool calls.

Prince Jha, Samuele Poppi, Nils Lukas · 0 citations
#machine learning Preprint Sep 2026

CopyShield: A Cross-Level Benchmark of Copyright Defenses in LLMs

Large language models can reproduce memorized text verbatim, yet copyright defenses are usually evaluated under incompatible protocols. We introduce CopyShield, a controlled benchmark comparing three representative defenses at distinct intervention levels: contrastive decoding (output), Direct Preference Optimization (...

Maryam Alshehyari, Dushyant Singh Chauhan, Samuele Poppi et al. · 0 citations
Preprint Aug 2026

ReACT-CLIP: Response-Aware Test-Time Defense for Vision--Language Models

This work introduces ReACT-CLIP, a response-conditioned test-time defense that separately determines how strongly each input should be corrected and whether defensive intervention is necessary, and quantifies this variation using a prediction-instability score computed by Jensen--Shannon divergence and combines it with...

H. Malik, Toluwani Aremu, Samuele Poppi et al. · 0 citations

A Gravitational Interpretation of Fine-Tuning Reversion

Fine-tuning on harmless data can partially undo behaviors acquired earlier in training. Safety can erode under benign post-alignment updates, unlearned capabilities can re-emerge, latent traits can transfer through apparently unrelated supervision, and related post-alignment fragility appears in other generative settin...

Samuele Poppi, Nils Lukas · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.