CoBind: Stage-Aware Compositional Binding for Training-Free Text-to-Image Generation
CoBind is introduced, a training-free framework for stage-aware compositional binding that parses a prompt into a composition graph of entities, attributes, and relations and adapts the guidance strength according to the current satisfaction of each constraint, reducing unnecessary latent updates.