Skip to content

Author

Tomasz Radzikowski

We have 1 of 1 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

DINO-A: Adapting Self-Distillation Vision Transformers to General Audio Representation Learning

DINO-A is presented, an adaptation of self-distillation from vision to general audio representation learning, and it is traced to two mechanisms: the interaction between DINO's high-dimensional projection space and FSD50K's limited scale, and the additional cost of multi-crop augmentation, which DINO uses but BYOL-A v2 does not.

Tomasz Radzikowski, M. Modrzejewski, Przemyslaw Rokita · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.