Zero-Shot Cross-Lingual Opinion Target Extraction

被引:0
作者
Jebbara, Soufian [1 ]
Cimiano, Philipp [2 ]
机构
[1] Semalytix GmbH, Bielefeld, Germany
[2] Bielefeld Univ, Semant Comp Grp, CITEC, Bielefeld, Germany
来源
2019 CONFERENCE OF THE NORTH AMERICAN CHAPTER OF THE ASSOCIATION FOR COMPUTATIONAL LINGUISTICS: HUMAN LANGUAGE TECHNOLOGIES (NAACL HLT 2019), VOL. 1 | 2019年
基金
欧盟地平线“2020”;
关键词
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Aspect-based sentiment analysis involves the recognition of so called opinion target expressions (OTEs). To automatically extract OTEs, supervised learning algorithms are usually employed which are trained on manually annotated corpora. The creation of these corpora is labor-intensive and sufficiently large datasets are therefore usually only available for a very narrow selection of languages and domains. In this work, we address the lack of available annotated data for specific languages by proposing a zero-shot cross-lingual approach for the extraction of opinion target expressions. We leverage multilingual word embeddings that share a common vector space across various languages and incorporate these into a convolutional neural network architecture for OTE extraction. Our experiments with 5 languages give promising results: We can successfully train a model on annotated data of a source language and perform accurate prediction on a target language without ever using any annotated samples in that target language. Depending on the source and target language pairs, we reach performances in a zero-shot regime of up to 77% of a model trained on target language data. Furthermore, we can increase this performance up to 87% of a baseline model trained on target language data by performing cross-lingual learning from multiple source languages.
引用
收藏
页码:2486 / 2495
页数:10
相关论文
共 23 条
[1]  
ALAM MS, 2016, SEMEVAL NAACL HLT, P1298
[2]  
Alvarez-Lopez T., 2016, P 10 INT WORKSHOP SE, P306
[3]  
Bojanowski Piotr, 2017, Trans. Assoc. Comput. Linguist., V5, P135, DOI DOI 10.1162/TACL_A_00051
[4]  
Caruana R, 2001, ADV NEUR IN, V13, P402
[5]   Aspect-Based Relational Sentiment Analysis Using a Stacked Neural Network Architecture [J].
Jebbara, Soufian ;
Cimiano, Philipp .
ECAI 2016: 22ND EUROPEAN CONFERENCE ON ARTIFICIAL INTELLIGENCE, 2016, 285 :1123-1131
[6]  
Kingma DP, 2014, ARXIV
[7]  
Lample G., 2018, 6 INT C LEARN REPR I
[8]  
Li Xin, 2017, EMNLP, P2886
[9]  
Nair V., 2010, P 27 INT C MACH LEAR, P807
[10]  
Ng A. Y., 2004, Proceedings of the twenty-first international conference on Machine learning, page, DOI [DOI 10.1145/1015330.1015435, 10.1145/1015330.1015435]