Preprint
Aug 2026
Language-Specific Gaps in AI Safety Training Datasets
It is connected to a documented, persistent asymmetry in multilingual jailbreak robustness (single-turn attacks largely mitigated, multi-turn attacks still effective), arguing that this asymmetry is structurally consistent with where the audit finds training and evaluation data thinnest.
Chialuka Prisca-Mary Onuoha, Bright Etornam Sunu, Rashidat Sikiru
· 0 citations