LG MLJul 30, 2020

Beyond $\mathcal{H}$-Divergence: Domain Adaptation Theory With Jensen-Shannon Divergence

Changjian Shui, Qi Chen, Jun Wen, Fan Zhou, Christian Gagné, Boyu Wang

arXiv:2007.15567v113.624 citations

Originality Highly original

AI Analysis

This addresses theoretical incoherence in domain adaptation for machine learning practitioners, providing a more flexible framework for various transfer learning scenarios.

The paper identifies a mismatch between domain adversarial training's optimization objective (Jensen-Shannon divergence) and its theoretical foundation (H-divergence), establishing new target risk bounds based on joint distributional Jensen-Shannon divergence. It demonstrates practical benefits with empirical validation on real datasets.

We reveal the incoherence between the widely-adopted empirical domain adversarial training and its generally-assumed theoretical counterpart based on $\mathcal{H}$-divergence. Concretely, we find that $\mathcal{H}$-divergence is not equivalent to Jensen-Shannon divergence, the optimization objective in domain adversarial training. To this end, we establish a new theoretical framework by directly proving the upper and lower target risk bounds based on joint distributional Jensen-Shannon divergence. We further derive bi-directional upper bounds for marginal and conditional shifts. Our framework exhibits inherent flexibilities for different transfer learning problems, which is usable for various scenarios where $\mathcal{H}$-divergence-based theory fails to adapt. From an algorithmic perspective, our theory enables a generic guideline unifying principles of semantic conditional matching, feature marginal matching, and label marginal shift correction. We employ algorithms for each principle and empirically validate the benefits of our framework on real datasets.

View on arXiv PDF

Similar