Skip to content

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Aug 2026

Sleight of Word Benchmark: Can Language Models Notice If Their Own Output Was Tampered With?

A simple benchmark is built in which a single word is consistently substituted with another in the generation process, and two distinct axes are measured: metrics that relate to the model's surprise, as well as an evaluation of the textual reaction for 19 different open-weight language models.

A. Cetoli · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.