Skip to content

Author

Gert Lek

We have 2 of 3 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Safety Reconstructed: Generative Modeling via Masked Diffusion Builds Strong Safety Guardrails

Guard models are the last line of defense between a language model and a harmful output, yet their training objective is surprisingly narrow. Existing guards learn to predict a single verdict token from a conversational context, concentrating supervision on a single target. The consequences are structural: models latch...

Gert Lek, Abele Malan, Chao-Yi Zhu et al. · 0 citations
Preprint Aug 2026

Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits

SN-Guided Diffusion is introduced, a fully offline black-box jailbreak framework that steers the diffusion process away from safety-triggering regions using a weighted safety neuron loss, which achieves near-perfect prompt separability.

Elena Dumitrescu, Gert Lek, L. Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.