Scalable and Accurate Self-supervised Multimodal Representation Learning without Aligned Video and Text Data.
Vladislav Lialin, Stephen Rawls, David Chan, Shalini Ghosh, Anna Rumshisky, Wael Hamza
Browse the full WACV paper archive.
Vladislav Lialin, Stephen Rawls, David Chan, Shalini Ghosh, Anna Rumshisky, Wael Hamza
Browse the full WACV paper archive.