Looking Similar, Sounding Different: Leveraging Counterfactual Cross-Modal Pairs for Audiovisual Representation Learning.
Nikhil Singh, Chih-Wei Wu, Iroro Orife, Mahdi M. Kalayeh
Browse the full CVPR paper archive.
Nikhil Singh, Chih-Wei Wu, Iroro Orife, Mahdi M. Kalayeh
Browse the full CVPR paper archive.