Distributional correspondence indexing for cross-lingual and cross-domain sentiment classification

Alejandro Moreo Fernández; Andrea Esuli; Fabrizio Sebastiani

Journal ArticleOPEN ACCESS

Distributional correspondence indexing for cross-lingual and cross-domain sentiment classification

Journal of Artificial Intelligence Research (2016) 55 131-163

DOI: 10.1613/jair.4762

34Citations

17Readers

Abstract

Domain Adaptation (DA) techniques aim at enabling machine learning methods learn effective classifiers for a "target" domain when the only available training data belongs to a different "source" domain. In this paper we present the Distributional Correspondence Indexing (DCI) method for domain adaptation in sentiment classification. DCI derives term representations in a vector space common to both domains where each dimension reects its distributional correspondence to a pivot, i.e., to a highly predictive term that behaves similarly across domains. Term correspondence is quantified by means of a distributional correspondence function (DCF). We propose a number of efficient DCFs that are motivated by the distributional hypothesis, i.e., the hypothesis according to which terms with similar meaning tend to have similar distributions in text. Experiments show that DCI obtains better performance than current state-of-the-art techniques for cross-lingual and cross-domain sentiment classification. DCI also brings about a significantly reduced computational cost, and requires a smaller amount of human intervention. As a final contribution, we discuss a more challenging formulation of the domain adaptation problem, in which both the cross-domain and cross-lingual dimensions are tackled simultaneously.

Cite

CITATION STYLE

APA

Fernández, A. M., Esuli, A., & Sebastiani, F. (2016). Distributional correspondence indexing for cross-lingual and cross-domain sentiment classification. Journal of Artificial Intelligence Research, 55, 131–163. https://doi.org/10.1613/jair.4762

Distributional correspondence indexing for cross-lingual and cross-domain sentiment classification

Abstract

Cite

Register to see more suggestions