Skip to main content

Weakly supervised cross-domain alignment with optimal transport

Publication ,  Journal Article
Yuan, S; Bai, K; Chen, L; Zhang, Y; Tao, C; Li, C; Wang, G; Henao, R; Carin, L
Published in: 31st British Machine Vision Conference, BMVC 2020
January 1, 2020

Cross-domain alignment between image objects and text sequences is key to many visual-language tasks, and it poses a fundamental challenge to both computer vision and natural language processing. This paper investigates a novel approach for the identification and optimization of fine-grained semantic similarities between image and text entities, under a weakly-supervised setup, improving performance over state-of-the-art solutions. Our method builds upon recent advances in optimal transport (OT) to resolve the cross-domain matching problem in a principled manner. Formulated as a drop-in regularizer, the proposed OT solution can be efficiently computed and used in combination with other existing approaches. We present empirical evidence to demonstrate the effectiveness of our approach, showing how it enables simpler model architectures to outperform or be comparable with more sophisticated designs on a range of vision-language tasks.

Duke Scholars

Published In

31st British Machine Vision Conference, BMVC 2020

Publication Date

January 1, 2020
 

Citation

APA
Chicago
ICMJE
MLA
NLM
Yuan, S., Bai, K., Chen, L., Zhang, Y., Tao, C., Li, C., … Carin, L. (2020). Weakly supervised cross-domain alignment with optimal transport. 31st British Machine Vision Conference, BMVC 2020.
Yuan, S., K. Bai, L. Chen, Y. Zhang, C. Tao, C. Li, G. Wang, R. Henao, and L. Carin. “Weakly supervised cross-domain alignment with optimal transport.” 31st British Machine Vision Conference, BMVC 2020, January 1, 2020.
Yuan S, Bai K, Chen L, Zhang Y, Tao C, Li C, et al. Weakly supervised cross-domain alignment with optimal transport. 31st British Machine Vision Conference, BMVC 2020. 2020 Jan 1;
Yuan, S., et al. “Weakly supervised cross-domain alignment with optimal transport.” 31st British Machine Vision Conference, BMVC 2020, Jan. 2020.
Yuan S, Bai K, Chen L, Zhang Y, Tao C, Li C, Wang G, Henao R, Carin L. Weakly supervised cross-domain alignment with optimal transport. 31st British Machine Vision Conference, BMVC 2020. 2020 Jan 1;

Published In

31st British Machine Vision Conference, BMVC 2020

Publication Date

January 1, 2020