Cross-Attention Head Position Patterns Can Align with Human Visual Concepts in Text-to-Image Generative Models.
Jungwon Park, Jungmin Ko, Dongnam Byun, Jangwon Suh, Wonjong Rhee
Browse the full ICLR paper archive.
Jungwon Park, Jungmin Ko, Dongnam Byun, Jangwon Suh, Wonjong Rhee
Browse the full ICLR paper archive.