Skip to content

Grounding Language Models to Images for Multimodal Inputs and Outputs.

Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried

VenueA*ICML
Year2023
ProceedingsICML

Browse the full ICML paper archive.