Skip to content

Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs.

Ling Yang, Zhaochen Yu, Chenlin Meng, Minkai Xu, Stefano Ermon, Bin Cui

VenueA*ICML
Year2024
ProceedingsICML

Browse the full ICML paper archive.