Wasserstein Distance Regularized Sequence Representation for Text Matching in Asymmetrical Domains
Weijie Yu, Chen Xu, Jun Xu, Liang Pang, Xiaopeng Gao, Xiaozhao Wang, Ji-Rong Wen
Abstract
One approach to matching texts from asymmetrical domains is projecting the input sequences into a common semantic space as feature vectors upon which the matching function can be readily defined and learned. In realworld matching practices, it is often observed that with the training goes on, the feature vectors projected from different domains tend to be indistinguishable. The phenomenon, however, is often overlooked in existing matching models. As a result, the feature vectors are constructed without any regularization, which inevitably increases the difficulty of learning the downstream matching functions. In this paper, we propose a novel match method tailored for text matching in asymmetrical domains, called WD-Match. In WD-Match, a Wasserstein distance-based regularizer is defined to regularize the features vectors projected from different domains. As a result, the method enforces the feature projection function to generate vectors such that those correspond to different domains cannot be easily discriminated. The training process of WD-Match amounts to a game that minimizes the matching loss regularized by the Wasserstein distance. WD-Match can be used to improve different text matching methods, by using the method as its underlying matching model. Four popular text matching methods have been exploited in the paper. Experimental results based on four publicly available benchmarks showed that WD-Match consistently outperformed the underlying methods and the baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1816af39-b87b-4087-a05e-1f8d782bb1fbCited by top-tier papers2
- Explainable Legal Case Matching via Inverse Optimal Transport-based Rationale ExtractionWeijie Yu, Zhongxiang Sun, Jun Xu, Zhenhua Dong et al.SIGIR 2022 · 45 citations
- Wasserstein Selective Transfer Learning for Cross-domain Text MiningLingyun Feng, Minghui Qiu, Yaliang Li, Haitao Zheng et al.EMNLP 2021 · 5 citations
Related papers
- RMIB: Representation Matching Information Bottleneck for Matching Text RepresentationsHaihui Pan, Zhifang Liao, Wenrui Xie, Kun HanICML 2024 · 1 citation
- Graph Optimal Transport for Cross-Domain AlignmentLiqun Chen, Zhe Gan, Yu Cheng, Linjie Li et al.ICML 2020 · 193 citations
- LAMDA: Label Matching Deep Domain AdaptationTrung Le, Tuan Nguyen, Nhat Ho, Hung Bui et al.ICML 2021 · 49 citations
- Normalized Wasserstein for Mixture Distributions With Applications in Adversarial Learning and Domain AdaptationYogesh Balaji, Rama Chellappa, Soheil FeiziICCV 2019 · 53 citations
- Wasserstein Transfer LearningKaicheng Zhang, Sinian Zhang, Doudou Zhou, Yidong ZhouNeurIPS 2025 · 2 citations
