Preprint
Aug 2026
MAVISEG: Manifold Propagation and Visual Prototypes for Zero-Shot Open-Vocabulary Segmentation in Diffusion Transformers
The results indicate that diffusion transformers carry more concept-level information than current attribution methods recover, and that much of it is lost on the way to the mask rather than absent from the model.
Rajatsubhra Chakraborty, Xujun Che, Ritabrata Chakraborty et al.
· 0 citations