Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Weakly supervised cross-domain alignment with optimal transport

Aug 14, 2020

Siyang Yuan, Ke Bai, Liqun Chen, Yizhe Zhang, Chenyang Tao, Chunyuan Li, Guoyin Wang, Ricardo Henao, Lawrence Carin

Figure 1 for Weakly supervised cross-domain alignment with optimal transport

Figure 2 for Weakly supervised cross-domain alignment with optimal transport

Figure 3 for Weakly supervised cross-domain alignment with optimal transport

Figure 4 for Weakly supervised cross-domain alignment with optimal transport

Share this with someone who'll enjoy it:

Abstract:Cross-domain alignment between image objects and text sequences is key to many visual-language tasks, and it poses a fundamental challenge to both computer vision and natural language processing. This paper investigates a novel approach for the identification and optimization of fine-grained semantic similarities between image and text entities, under a weakly-supervised setup, improving performance over state-of-the-art solutions. Our method builds upon recent advances in optimal transport (OT) to resolve the cross-domain matching problem in a principled manner. Formulated as a drop-in regularizer, the proposed OT solution can be efficiently computed and used in combination with other existing approaches. We present empirical evidence to demonstrate the effectiveness of our approach, showing how it enables simpler model architectures to outperform or be comparable with more sophisticated designs on a range of vision-language tasks.

* Accepted to BMVC 2020 (Oral)

View paper on

Share this with someone who'll enjoy it:

Title:Weakly supervised cross-domain alignment with optimal transport

Paper and Code