Business Faculty Articles and Research

Ideology Prediction from Scarce and Biased Supervision: Learn to Disregard the “What” and Focus on the “How”!

Chen Chen, CUHK ShenzhenFollow
Dylan Walker, Chapman UniversityFollow
Venkatesh Saligrama, Boston UniversityFollow

Document Type

Conference Proceeding

Publication Date

7-2023

Abstract

We propose a novel supervised learning approach for political ideology prediction (PIP) that is capable of predicting out-of-distribution inputs. This problem is motivated by the fact that manual data-labeling is expensive, while self-reported labels are often scarce and exhibit significant selection bias. We propose a novel statistical model that decomposes the document embeddings into a linear superposition of two vectors; a latent neutral context vector independent of ideology, and a latent position vector aligned with ideology. We train an end-to-end model that has intermediate contextual and positional vectors as outputs. At deployment time, our model predicts labels for input documents by exclusively leveraging the predicted positional vectors. On two benchmark datasets we show that our model is capable of outputting predictions even when trained with as little as 5% biased data, and is significantly more accurate than the state-of-the-art. Through crowd-sourcing we validate the neutrality of contextual vectors, and show that context filtering results in ideological concentration, allowing for prediction on out-of-distribution examples.

Comments

This article was originally published in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics, volume 1, in 2023. https://doi.org/10.18653/v1/2023.acl-long.530

Recommended Citation

Chen Chen, Dylan Walker, and Venkatesh Saligrama. 2023. Ideology Prediction from Scarce and Biased Supervision: Learn to Disregard the “What” and Focus on the “How”!. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 9529–9549, Toronto, Canada. Association for Computational Linguistics. https://doi.org/10.18653/v1/2023.acl-long.530

Copyright

Association for Computational Linguistics

Creative Commons License

This work is licensed under a Creative Commons Attribution 4.0 License.

Download

Included in

Computational Linguistics Commons, Data Science Commons

COinS

Chapman University Digital Commons

Business Faculty Articles and Research

Ideology Prediction from Scarce and Biased Supervision: Learn to Disregard the “What” and Focus on the “How”!

Document Type

Publication Date

Abstract

Comments

Recommended Citation

Copyright

Creative Commons License

Included in

Browse

Search

Author Corner

Links

Chapman University Digital Commons

Business Faculty Articles and Research

Ideology Prediction from Scarce and Biased Supervision: Learn to Disregard the “What” and Focus on the “How”!

Authors

Document Type

Publication Date

Abstract

Comments

Recommended Citation

Copyright

Creative Commons License

Included in

Share

Browse

Search

Author Corner

Links