Preprint
Aug 2026
FlowSep 2: Self-Supervised Flow Matching for Language-Queried Audio Source Separation
This work proposes FlowSep2, a text-conditioned flow-matching generative model for LASS, which learns to generate the target source representation from Gaussian noise in a latent space, conditioned on both the mixture representation and the text query.
Yiitan Yuan, Xubo Liu, Haohe Liu et al.
· 0 citations