Publication

CommunityLM: Probing Partisan Worldviews from Language Models

hjian42

Sept. 28, 2022

Topics

People

Projects

CommunityLM

Groups

Share this publication

Hang Jiang, Doug Beeferman, Brandon Roy, and Deb Roy. 2022. CommunityLM: Probing Partisan Worldviews from Language Models. In Proceedings of the 29th International Conference on Computational Linguistics (COLING), pages 6818–6826, Gyeongju, Republic of Korea. International Committee on Computational Linguistics.

Abstract

As political attitudes have diverged ideologically in the United States, political speech has diverged lingusitically. The ever-widening polarization between the US political parties is accelerated by an erosion of mutual understanding between them. We aim to make these communities more comprehensible to each other with a framework that probes community-specific responses to the same survey questions using community language models CommunityLM. In our framework we identify committed partisan members for each community on Twitter and fine-tune LMs on the tweets authored by them. We then assess the worldviews of the two groups using prompt-based probing of their corresponding LMs, with prompts that elicit opinions about public figures and groups surveyed by the American National Election Studies (ANES) 2020 Exploratory Testing Survey. We compare the responses generated by the LMs to the ANES survey results, and find a level of alignment that greatly exceeds several baseline methods. Our work aims to show that we can use community LMs to query the worldview of any group of people given a sufficiently large sample of their social media discussions or media diet. Our code is publicly available: https://github.com/hjian42/communitylm.

via International Committee on Computational Linguistics

Using Twitter Data to Understand Public Perceptions of Approved versus Off-label Use for COVID-19-related Medications

Hua, Yining*, Hang Jiang*, Shixu Lin, Jie Yang, Joseph M. Plasek, David W. Bates, and Li Zhou. "Using Twitter Data to Understand Public Perceptions of Approved versus Off-label Use for COVID-19-related Medications." Journal of the American Medical Informatics Association (2022). *Equal Contribution.

Publication Research

CommunityLM: Probing Partisan Worldviews from Language Models

Topics

People

Projects

Groups

Abstract

Using Twitter Data to Understand Public Perceptions of Approved versus Off-label Use for COVID-19-related Medications

LNN-EL: A Neuro-Symbolic Approach to Short-text Entity Linking

The Birth of a Word

Human-Machine Collaboration for Rapid Speech Transcription

CommunityLM: Probing Partisan Worldviews from Language Models

Topics

People

Projects

Groups

Share this publication

Abstract

Using Twitter Data to Understand Public Perceptions of Approved versus Off-label Use for COVID-19-related Medications

LNN-EL: A Neuro-Symbolic Approach to Short-text Entity Linking

The Birth of a Word

Human-Machine Collaboration for Rapid Speech Transcription