Detecting political biases of named entities and hashtags on Twitter

被引:3
作者
Xiao, Zhiping [1 ]
Zhu, Jeffrey [1 ]
Wang, Yining [1 ]
Zhou, Pei [2 ]
Lam, Wen Hong [1 ]
Porter, Mason A. [3 ,4 ]
Sun, Yizhou [1 ]
机构
[1] Univ Calif Los Angeles, Dept Comp Sci, 580 Portola Pl, Los Angeles, CA 90095 USA
[2] Univ Southern Calif, Informat Sci Inst, Marina del Rey, Los Angeles, CA 90292 USA
[3] Univ Calif Los Angeles, Dept Math, 520 Portola Pl, Los Angeles, CA 90095 USA
[4] Santa Fe Inst, 1399 Hyde Pk Rd, Santa Fe, NM 87501 USA
基金
美国国家航空航天局; 美国国家科学基金会;
关键词
Political-polarity detection; Word embeddings; Multi-task learning; Adversarial training; Data sets; SENTIMENT ANALYSIS;
D O I
10.1140/epjds/s13688-023-00386-6
中图分类号
O1 [数学];
学科分类号
0701 ; 070101 ;
摘要
Ideological divisions in the United States have become increasingly prominent in daily communication. Accordingly, there has been much research on political polarization, including many recent efforts that take a computational perspective. By detecting political biases in a text document, one can attempt to discern and describe its polarity. Intuitively, the named entities (i.e., the nouns and the phrases that act as nouns) and hashtags in text often carry information about political views. For example, people who use the term "pro-choice" are likely to be liberal and people who use the term "pro-life" are likely to be conservative. In this paper, we seek to reveal political polarities in social-media text data and to quantify these polarities by explicitly assigning a polarity score to entities and hashtags. Although this idea is straightforward, it is difficult to perform such inference in a trustworthy quantitative way. Key challenges include the small number of known labels, the continuous spectrum of political views, and the preservation of both a polarity score and a polarity-neutral semantic meaning in an embedding vector of words. To attempt to overcome these challenges, we propose the Polarity-aware Embedding Multi-task learning (PEM) model. This model consists of (1) a self-supervised context-preservation task, (2) an attention-based tweet-level polarity-inference task, and (3) an adversarial learning task that promotes independence between an embedding's polarity component and its semantic component. Our experimental results demonstrate that our PEM model can successfully learn polarity-aware embeddings that perform well at tweet-level and account-level classification tasks. We examine a variety of applications-including a study of spatial and temporal distributions of polarities and a comparison between tweets from Twitter and posts from Parler-and we thereby demonstrate the effectiveness of our PEM model. We also discuss important limitations of our work and encourage caution when applying the PEM model to real-world scenarios.
引用
收藏
页数:26
相关论文
共 59 条
[1]   Exposure to opposing views on social media can increase political polarization [J].
Bail, Christopher A. ;
Argyle, Lisa P. ;
Brown, Taylor W. ;
Bumpus, John P. ;
Chen, Haohan ;
Hunzaker, M. B. Fallin ;
Lee, Jaemin ;
Mann, Marcus ;
Merhout, Friedolin ;
Volfovsky, Alexander .
PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2018, 115 (37) :9216-9221
[2]  
Barbera P, 2015, PREPRINT
[3]  
Batra S., 2010, Entity Based Sentiment Analysis on Twitter
[4]  
Bengio Y, 2001, ADV NEUR IN, V13, P932
[5]   The new Voteview.com: preserving and continuing Keith Poole's infrastructure for scholars, students and observers of Congress [J].
Boche, Adam ;
Lewis, Jeffrey B. ;
Rudkin, Aaron ;
Sonnet, Luke .
PUBLIC CHOICE, 2018, 176 (1-2) :17-32
[6]  
Bose AJ., 2019, P 36 INT C MACH LEAR
[7]  
Cawley GC, 2010, J MACH LEARN RES, V11, P2079
[8]  
Chao ZH, 2022, Arxiv, DOI arXiv:2212.00237
[9]   #Election2020: the first public Twitter dataset on the 2020 US Presidential election [J].
Chen, Emily ;
Deb, Ashok ;
Ferrara, Emilio .
JOURNAL OF COMPUTATIONAL SOCIAL SCIENCE, 2022, 5 (01) :1-18
[10]  
Chen Xi, 2016, Advances in Neural Information Processing Systems, V29